マルチモーダル A
125 件中 121〜125 件目を表示
-
NVIDIA Ising Enables Fully Automated Quantum Computer Calibration with Enhanced In-Context LearningNVIDIA、量子コンピュータの較正を自動化するOSS視覚言語モデルを公開NVIDIAは、量子プロセッサの診断出力を解釈し較正手順を判断するオープンソースの視覚言語モデル(VLM)「Ising Calibration」を公開した。文脈内学習(in-context learning)を強化し、量子コンピュータの較正を完全自動化することを狙う。※本文抜粋が途中までのため、詳細な性能や対応機種は確認できない。
-
CADER: Confidence-Aware Dynamic Evidence Reasoning for Long-Video Understanding
-
The Visual Bottleneck: Sparse-Frame Adaptation of MLLMs for Joint Spatial-Temporal Video Grounding
-
EchoBridge: Long-Tail-Aware ECG-Echocardiography Text Alignment for Echocardiography-Derived Cardiac Findings
-
Task-Conditional Faithfulness Auditing of Multimodal LLMs for Grid Diagnosis