Multimodal– tag –
-
New Technology
Multimodal LLM Agent ‘SynAgent’ Achieves Autonomous Synthesis of Highly Crystalline LiCoO2 Thin Films and Elucidates Substrate Temperature Control in Just 18 Experiments
arXiv USA Overview The 'SynAgent' framework, utilizing multimodal LLM agents, successfully achieved autonomous synthesis of highly crystalline LiCoO2 (001) thin films and elucidated the role of substrate temperature in crystallization wi... -
New Technology
Seoul National University AI Discovers Two Lead-Free High-k Dielectric Materials from 150 Million Virtual Compositions for Future Electronics
Almerja South Korea Overview Researchers at Seoul National University leveraged AI to screen approximately 150 million virtual chemical compositions, identifying two promising lead-free dielectric materials for future electronics. This '... -
New Technology
Google Gemini Omni Flash, ByteDance Seedance 2.0, Kuaishou Kling 3.0 Lead 2026 AI Video Generation Models
AI Hub Global Overview In 2026, Google's Gemini Omni Flash dominates the leaderboard for both text-to-video and image-to-video AI generation. ByteDance's Seedance 2.0 is hailed as the overall best AI video generation model, recognized fo... -
New Technology
OpenRouter’s September 2026 Video Generation Model Ranking: Seedance Series and Alibaba Wan 3.0 Lead the Pack
OpenRouter Global Overview OpenRouter released its updated September 2026 ranking of video generation models, featuring Seedance 2.0 Mini, Seedance 2.5, and Veo 3.1 Lite as top contenders. Notably, Alibaba's Wan 3.0 stands out for its ve... -
New Technology
Open-Source LLMs GLM-5.2, Llama 4 Maverick, and Kimi K3 Lead Benchmarks in September 2026
Thunder Compute Global Overview In the September 2026 open-source LLM rankings, GLM-5.2, Llama 4 Maverick, and Kimi K3 demonstrated leading performance across various domains. The 744B MoE GLM-5.2 (40B parameters) excelled in reasoning b... -
New Technology
ACS Publications: PolyCLIP Framework Outperforms Existing Polymer Property Prediction Models by 13.32%, Establishes New Foundation for Multimodal Polymer Informatics
ACS Publications USA Overview ACS Publications announced PolyCLIP, a CLIP-based multimodal framework, achieved up to 13.32% better performance in polymer property prediction than existing unimodal and multimodal models. The study demonst... -
New Technology
SciTechDaily: AI Discovers Two Promising High-Temperature Dielectric Candidates from 150 Million Virtual Materials for Future Electronics, Meeting EV and Aerospace Demands
SciTechDaily USA Overview SciTechDaily reports that an AI-driven method identified two promising dielectric material candidates from 150 million virtual chemical spaces for future electronics. This inverse design strategy combines multim... -
New Technology
Synaptics Launches Torq Edge AI Platform Featuring Google’s RISC-V Coral Open NPU, Boosting Multimodal Inference for Embedded IoT
Edge AI and Vision Insights USA Overview Synaptics unveiled its "Torq™ Edge AI platform," a high-performance, low-power Neural Processing Unit (NPU) subsystem for edge AI, integrated with Google's RISC-V based Coral Open NPU. Part of the... -
New Technology
September 2026 LLM Benchmark Update: Anthropic’s Claude Fable 5.1 Achieves 82.589%, Google’s Gemini 3.5 Flash Scores 83.6% on MMMU
LLM Gateway, LLM Stats International Overview The latest September 2026 LLM leaderboard reveals significant performance gains across multiple AI models in reasoning, coding, and vision tasks. Anthropic's Claude Fable 5.1 achieved an over... -
New Technology
Google DeepMind Unveils Gemini Ultra 1.5: 2 Million Token Context Window and Reduced Inference Costs Drive Multimodal Reasoning Breakthrough
Google DeepMind Official Blog USA Overview Google DeepMind has launched Gemini Ultra 1.5, significantly enhancing multimodal reasoning across text, image, and video, demonstrating state-of-the-art performance on new complex reasoning ben...