Skip to main content
株式会社オブライト

Articles tagged "マルチモーダル"

5 articles

AI2026-08-22
DeepSeek-V4-Flash-Vision-Exp: A First Look at the New Multimodal API
DeepSeek's V4-Flash-Vision-Exp launched Aug 21, 2026, adds image input at no premium over V4-Flash. Images run 384 tokens max; Files API free. Pricing, usage.
DeepSeek V4MoEマルチモーダル
AI2026-08-03
Qwen3.7 Flash Explained: $0.03 per Million Tokens, 1M Context, and What This Closed Vision-Language Model Can (and Can't) Do (2026)
Qwen3.7 Flash: Alibaba's closed vision model, priced from $0.03 per million input tokens, 1M-token context. Tiered pricing, use cases, limits. Updated Aug. 2026.
Qwen 3.7AlibabaAPI料金
AI2026-04-17
Claude Opus 4.7 Complete Guide — SWE-bench 87.6%, Vision 98.5% & New xhigh Effort Mode [April 16, 2026 Release]
Released April 16, 2026, Claude Opus 4.7 achieves SWE-bench Verified 87.6%, Vision accuracy 98.5%, and introduces the new xhigh Effort Control — all at the same price as Opus 4.6. This guide covers every major upgrade to Anthropic's latest flagship model.
Claude Opus 4.7AnthropicSWE-bench
AI2026-04-10
Mistral Small 4 Complete Guide — Unified Reasoning, Multimodal & Code in 119B MoE [2026]
Mistral Small 4, released March 2026, unifies reasoning, multimodal vision, and agentic coding in a 119B MoE model under Apache 2.0. Supports 11 languages including Japanese. Full specs, setup guide, and model comparisons.
Mistral Small 4MoEマルチモーダル
AI2026-03-04
Qwen3.5-9B Multimodal Guide: Running Free Image & Video AI In-House
Learn how to leverage Qwen3.5-9B's early-fusion multimodal architecture for free in-house image and video AI. Covers OCR, product inspection, surveillance analysis, meeting summarization, cloud API comparison, and step-by-step setup for local multimodal inference.
Qwen3.5マルチモーダル画像認識