AI News List

List of AI News about multimodal

Time Details
2026-08-15
15:00
PixVerse Elevates Emotional Video AI

According to PixVerse_ on X, creators can direct emotion in AI video, signaling controllable performance gains in text to video tools.

Source
2026-08-07
15:00
PixVerse Reveals global filmmaking push

According to PixVerse... The company teases an August 7 launch to grow AI filmmaking from clips to persistent characters and worlds, per its X post.

Source
2026-08-06
11:50
Mootion 5.0 Delivers Drama-Style Video Breakthrough

According to Mootion_AI, Mootion 5.0 generated a show-style drama clip and invites users to guess the reference for a prompt breakdown.

Source
2026-08-04
17:35
FLUX 3 Debuts action prediction breakthrough

According to Krea... FLUX 3 adds real‑world action prediction beyond video and launches on Krea, with more modalities coming soon, per Krea on X.

Source
2026-08-04
04:38
PixVerse Showcases AI Video Breakthrough

According to PixVerse... A viral 2026 space-odyssey style clip spotlights AI video generation’s creative speed and low-cost workflows.

Source
2026-08-02
03:00
Claude Opus 5 renders LoTR scene in Three.js

According to @karpathy, Claude Opus 5 spent ~2 hours and 1M tokens to generate 5,500 lines of Three.js code rendering a LotR scene.

Source
2026-08-02
03:00
Claude Opus 5 renders LoTR in Three.js

According to karpathy, Claude Opus 5 used a 1M token run to generate 5,500 lines of Three.js code rendering LoTR, revealing promise and audit limits.

Source
2026-07-31
15:03
MiniMax H3 powers PixVerse 2K video workflow

According to PixVerse... H3 adds 2K video, native audio, and multimodal refs, plus limited-time member discounts and credit promo.

Source
2026-07-30
17:55
Inkling Small Debuts: 4x Smaller, Big Results

According to soumithchintala, Inkling-Small matches Inkling quality at 4x smaller with 276B params, 12B active, and open weights for fine-tuning.

Source
2026-07-30
16:00
Gemini macOS Voice Boosts Workflow

According to @GeminiApp, Gemini for macOS adds voice to dictate, transform text, analyze files, and create visuals in any app, rolling out globally in English.

Source
2026-07-24
04:34
Gemini Data Reveals Multimodal Upside

According to @emollick, Google shares Gemini usage data showing multimodal AI benefits manual labor more than expected, per Google’s research blog.

Source
2026-07-19
17:11
Gemma 4 Visualizes multimodal thinking

According to emollick, a video shows Gemma 4 12B’s predicted next-token outputs over raw image patches, revealing multimodal attention dynamics.

Source
2026-07-16
23:41
Kimi K3 Debuts with 2.8T Params, 1M Context

According to soumithchintala, Kimi K3 ships 2.8T params, 1M context, multimodal, faster decoding, and will open weights by July 27, 2026, per Kimi.ai.

Source
2026-07-16
01:58
Seedance 2.5 Demos Stun: ByteDance Breakthrough

According to The Rundown AI, ByteDance’s Seedance 2.5 shows striking video generation quality and fluid motion in new demo clips.

Source
2026-07-15
19:16
Inkling Multimodal Model Debuts with 1M Context

According to TheRundownAI, Thinking Machines launched Inkling, an open-weight multimodal model with 1M tokens and full weights on Hugging Face.

Source
2026-07-15
18:46
ChatGPT 5.6 Sol Pro cracks clue‑less crossword

According to emollick, ChatGPT 5.6 Sol Pro filled a clue‑less Pokémon crossword, signaling rapid multimodal reasoning gains for enterprise use.

Source
2026-07-15
18:15
Inkling Launches 975B Open Weights Multimodal Model

According to soumithchintala, Inkling debuts with 975B params, open weights, and native text image audio support on Tinker and Hugging Face.

Source
2026-07-15
14:59
Claude3.5 Showcases nuanced reasoning in new video

According to TheRundownAI, Claude highlights nuanced reasoning and hope in hard questions in a new video, signaling richer enterprise use cases.

Source
2026-07-15
01:32
Meta Muse Spark 1.1 Powers OpenClaw Upgrade

According to @openclaw, Meta’s Muse Spark 1.1 is now live in OpenClaw v2026.7.1, enabling multimodal reasoning for agentic coding and tool use.

Source
2026-07-14
21:09
Meta AI model aces APhO exam with perfect 30

According to AIatMeta, Meta’s model scored 30/30 on the Asian Physics Olympiad theoretical exam, tying top 3 students, showcasing advanced multimodal reasoning.

Source