Multimodal– tag –
-
New Technology
Multimodal Foundation Model SQUALL Integrates Histology with Spatial Molecular Programs, Revolutionizing Cancer Biomarker Profiling and Outcome Prediction
bioRxiv International Overview Researchers have developed SQUALL, a multimodal foundation model that integrates histology with spatial molecular programs to deepen the mechanistic interpretation of histopathological assessments. Pretrain... -
New Technology
Tempus AI Unveils Promising Initial Results from Oncology Multimodal Foundation Model Trained on 2.5M Longitudinal Records, Achieving C-index of 0.802 for OS Prediction
Business Wire USA Overview Tempus AI announced initial results at ASCO 2026 for its multimodal foundation model aimed at oncology insight generation. This transformer-based model, trained on 2.5 million longitudinal records, 450,000 digi... -
New Technology
FDA-Cleared AI “Artera AI Breast” Demonstrates Clinical Utility for Breast Cancer Prognosis and Chemotherapy Benefit Prediction at ASCO 2026
The Pathologist USA Overview At ASCO 2026, a study showcased the utility of Artera AI Breast, an FDA-cleared multimodal artificial intelligence model, for predicting prognosis and chemotherapy benefit in postmenopausal women with node-po... -
New Technology
ImmunoFoundation: Novel Multimodal Foundation Model for Immunogenicity Prediction and Peptide Optimization Overcomes Data Scarcity
OpenReview USA Overview Researchers have introduced ImmunoFoundation, a self-supervised multimodal foundation model designed for immunogenicity prediction and peptide optimization. By integrating an ESM-2 sequence encoder with a graph tr... -
New Technology
Anthropic Launches Claude Opus 4.8: Enhanced Reasoning and Extended Context Window
LLM Stats USA Overview Anthropic released its latest proprietary large language model, Claude Opus 4.8, on May 27, 2026, marking a significant upgrade in its top-tier Opus series. This new iteration promises enhanced reasoning capabiliti... -
New Technology
Advanced LLMs of 2026: Benchmarking Top Models for Reasoning, Coding, and Multimodal Capabilities
aimlapi.com USA Overview As of May 2026, the LLM landscape is dominated by models such as GPT-5.5, Claude Opus 4.7, and DeepSeek V4 Pro, characterized by the mainstreaming of agent architectures and the standardization of 1-million-token... -
New Technology
Google Unveils Gemini Omni: A New AI World Model with Advanced Physically Accurate Video Generation
(unspecified tech media, possibly Mashable Light Speed) USA Overview At Google I/O 2026, Google introduced Gemini Omni, its new AI world model. This multimodal model accepts text, audio, image, and video inputs, leveraging Gemini's "real... -
New Technology
Google I/O 2026 to Unveil Next-Gen Gemini and ‘Omni’ Video Generator, Intensifying AI Competition
(unspecified tech media) USA Overview AI is set to dominate Google I/O 2026 with the anticipated release of a new Gemini version and the unannounced 'Omni' video generation model. Leaks suggest Gemini Omni will enable direct video creati... -
New Technology
May 2026 AI Update: GPT-5.5 Instant Reduces Hallucinations by 52%, Enhances Personal Context
LLM Stats USA Overview In early May 2026, xAI released Grok 4.3 and OpenAI launched GPT-5.5 Instant, now the default for ChatGPT. GPT-5.5 Instant significantly reduced hallucinations by 52% and improved accuracy in multimodal tasks like ... -
New Technology
NVIDIA Unveils Neotron 3 Nano Omni: A Unified Multimodal AI for Efficient Agent Deployment
MindStudio USA Overview NVIDIA has launched Neotron 3 Nano Omni, an open-weight multimodal AI model capable of processing text, images, video, and audio within a single, efficient architecture. Optimized for real-world deployment on acce...