Search

Tag: #multimodal-llm14 results

Huawei HarmonySpace 6 with Chatty AI Agent

Huawei HarmonySpace 6 with Chatty AI Agent

Huawei launched HarmonySpace 6 featuring Xiao Yi intelligent agent, industry's first full-scene chatting AI for navigation, control, and services while driving. Upgrades include AMS multimodal perception for vital signs and safety, MoLA 2.0 with 100B+ param multimodal model and Agent platform. New hardware like dual 17.2-inch screens, million-pixel lights, and laser projector enhance entertainment.

Xiaomi SU7 Ships with XLA Cognitive Model

Xiaomi SU7 Ships with XLA Cognitive Model

New-generation Xiaomi SU7 launches fully equipped with Xiaomi XLA cognitive large model for advanced assisted driving, featuring top-tier hardware like 700TOPS Thor chip and multiple sensors. XLA enables multimodal fusion including voice-controlled driving and parking. OTA upgrades will roll out to first-gen SU7 Pro/Max/Ultra and YU7 models.

RideJudge: AI Framework for Ride Disputes

RideJudge: AI Framework for Ride Disputes

RideJudge introduces a progressive framework aligning visual and logical reasoning for ride-hailing dispute adjudication, addressing limitations in multimodal LLMs. Key innovations include SynTraj for trajectory synthesis, Adaptive Context Optimization, and Ordinal-Sensitive RL. The 8B model achieves 88.41% accuracy, outperforming 32B baselines.

PlotChain Benchmark for MLLM Plot Reading

PlotChain Benchmark for MLLM Plot Reading

PlotChain introduces a deterministic benchmark for evaluating multimodal LLMs on extracting quantitative values from engineering plots like Bode and FFT. It features 450 plots across 15 families with ground truth and checkpoint diagnostics for failure analysis. Top models score ~80% (Gemini 2.5 Pro leads), but frequency tasks remain weak.

ArXiv AIResearchFeb 17#multimodal-llm#engineering-plots
Page 1 of 2