New Technology– category –
-
New Technology
Google Unveils Gemini 3.8 Live Models, Revolutionizing Real-Time Voice AI with Multilingual Multistep Reasoning
TechShots USA Overview Google has launched new audio models, Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, significantly advancing real-time voice AI capabilities. These models can execute tools and API calls mid-conversation, u... -
New Technology
Creatify Labs Launches Boreal, a Text-to-Video AI Model Delivering Top-Tier Quality at One Cent Per Second and 40x Speed
PRWeb United States Overview Creatify Labs has unveiled Boreal, a groundbreaking AI video model capable of generating video from text and images in real-time. Boreal delivers video quality comparable to top-tier closed models, but at a s... -
New Technology
Tsinghua University and Shengshu Technology Unveil ‘Vidu S2’ Real-Time Interactive, Editable, and Spatial Video Generation Model, Boosting 720p Generation and Instruction Following
arXiv (Tsinghua University, Shengshu Technology) China Overview Tsinghua University and Shengshu Technology researchers have announced 'Vidu S2,' a real-time interactive, editable, and spatial video generation model on arXiv. Comprising ... -
New Technology
Danube Dynamics and EVVA Leverage AI to Slash High-Reflective Component Surface Defect Inspection Time from 30s to Under 5s, Boosting Productivity by 6x
unconfirmed (via Metrology News) Austria Overview A collaboration between Austria's Danube Dynamics and EVVA has yielded an AI-based inspection system for surface defects on highly reflective components. This system automatically evaluat... -
New Technology
University of Surrey and NVIDIA Develop New Training Method to Enhance AI-Generated Scene Responsiveness to User Camera Controls
University of Surrey and NVIDIA (via unconfirmed preprint) United Kingdom Overview Researchers from the University of Surrey and NVIDIA have developed a novel training method that enables AI-generated scenes to respond more accurately to... -
New Technology
New LLM Inference Algorithm for Compute-in-Flash Systems Achieves 15x KV Cache Traffic Reduction, Delivering Energy and Latency Savings for Llama-3.1-8B and Qwen-2.5-7B
arXiv Unknown Overview Research published on arXiv introduces a novel algorithm enabling LLM inference on Compute-in-Flash systems, mitigating memory bandwidth limitations. This method employs end-to-end integer-only quantization and a d... -
New Technology
New ‘TAM’ Benchmark Exposes Critical Gaps in GPT-5’s Long-Horizon Procedural Reasoning, Achieving Only 1% Exact Match on ICD-10-CM Clinical Coding
arXiv Unknown Overview A new benchmark, Tasks over Application Manuals (TAM), has been introduced on arXiv to accurately assess LLMs' long-horizon procedural reasoning. Evaluating GPT-5, TAM revealed alarmingly low exact match performanc... -
New Technology
LLM Test-Time Scaling: Candidate Generation Strategy Dramatically Increases Energy Consumption by 4.86x and Latency by 6.12x on A100 GPUs for Phi-3-mini and Qwen2.5-1.5B
Mobina Kashaniyan (referencing IEEE/ACM SC26 Workshop) Unknown Overview New research reveals that for LLM test-time scaling, the candidate generation strategy, not just the candidate count, critically impacts energy and performance. Eval... -
New Technology
LLMAR: Tuning-Free Framework Boosts Recommendation Performance in Sparse, Text-Rich Industrial Domains via LLM Inference and Self-Verification
arXiv Unknown Overview arXiv introduces LLMAR, a tuning-free recommendation framework for sparse and text-rich industrial domains, demonstrating superior accuracy, explainability, and operational cost efficiency compared to traditional t... -
New Technology
arXiv Introduces EgoPathBench: New Benchmark Reveals Limits of Zero-Shot Egocentric Waypoint Decision-Making in Vision-Language Models
arXiv Unknown Overview A new benchmark, EgoPathBench, has been introduced on arXiv to evaluate the zero-shot egocentric waypoint decision-making capabilities of Vision-Language Models (VLMs). Comprising 31,852 training, 1,345 validation,...