It's Chinese New Year and I'm about to leave home. So I asked @GeminiApp to compose a farewell track — folk-pop, shifting from festive warmth to quiet solitude.
Seconds later, Lyria 3 delivered a 30-second piece that genuinely moved me.
30 seconds is short. But when AI captures emotion this well, it hits different. 🎵
Introducing Lyria 3, our latest and most advanced music model, available in the Gemini App starting today : )
Go from idea, image, or video to music in seconds!
🎙️Ming-omni-tts-16.8B-A3B, Ming-omni-tts-0.5B, are the Voice Core of Ming-flash-omni-2.0, ready for:
-Short-video creator looking for the perfect voiceover
-Podcaster turns a monologe into a dynamic duo
-Dev looking to integrate a voice assistant into your OpenClaw project
One for all, and all for one 🧧
Introducing Ming-flash-omni-2.0: A specialist in every domain, unified as a capable generalist. A gift from Ling =)
- Unified Acoustic Synthesis: Speech, audio, and music combined for unbounded creativity;
- "Seeing" to "Knowing": Moving beyond input to true deep semantic understanding;
- Native Visual Fusion: Seamless generation, editing, and segmentation;
(2/3) Context-Aware ASR and Superior Dialect Recognition: The model achieves full SOTA across all 12 ContextASR subtasks. Its ability to understand and process 15 distinct Chinese dialects (e.g., Hunanese) is substantially improved. This advanced audio comprehension offers critical real-time translation and interpretation support, aiding users navigating unfamiliar linguistic contexts. #LLM#SpeechRecognition 🌍
(4/4) Enhanced Voice Cloning: The voice generation capability is significantly upgraded through a transition from discrete to continuous tokenizers. This advancement dramatically improves voice cloning accuracy and ensures high stability for both Mandarin and English mixed-language dialogue. The model can effectively clone the original speaker's timbre into new dialogue, achieving a remarkable seed-tts-zh WER score of 0.99, surpassing Qwen3 Omni and Seed-TTS. #VoiceCloning
🚀We officially release Ring-1T, the open-source trillion-parameter thinking model built on the Ling 2.0 architecture.
Ring-1T achieves silver-level IMO reasoning through pure natural language reasoning.
→ 1 T total / 50 B active params · 128 K context window
→ Reinforced by Icepop RL + ASystem (Trillion-Scale RL Engine)
→ Open-source SOTA in natural language reasoning — AIME 25 / HMMT 25 / ARC-AGI-1 / CodeForce
Deep thinking · Open weights · FP8 version available
We're still training and pushing to do more. Stay tuned!
1/5
🚀 Ling-1T — Trillion-Scale Efficient Reasoner
Introducing Ling-1T, the first flagship non-thinking model in the Ling 2.0 series —
1 Trillion total parameters with ≈ 50 B active per token, trained on 20 T+ reasoning-dense tokens.
Highlights
→ Evo-CoT curriculum + Linguistics-Unit RL for scalable reasoning
→ Strong efficiency–accuracy balance on complex reasoning tasks
→ Advanced visual understanding + front-end code generation via Syntax–Function–Aesthetics reward
→ Emergent tool-use ability (≈ 70 %) with minimal instruction tuning
→ FP8 mixed-precision + Ling Scaling Law → efficient trillion-scale training
Efficient Thinking · Precise Reasoning
Ling-1T extends the Pareto frontier of reasoning accuracy vs. cost —
a new milestone in open-source trillion-scale intelligence.
🫡 every week we treat your fomo with open AI release highlights:
> new Qwen3-VL models (3B/30B MoE) better than GPT-5-Mini 🔥
> two new DeepSeek models (v3.2 experimental)
> Ovi is a new text-to-video model with sounds ⏯️
> new Ming image and audio editing models
find more ⤵️
Here's my 4+ hour conversation with Pavel Durov (@durov), founder and CEO of Telegram. This was one of the most fascinating and powerful conversations I've ever had in my life.
We discuss everything from his philosophy on freedom to government bureaucracies, intelligence
222K Followers 3K FollowingFollow for posts about GitHub repos, DSPy, and agents
Subscribe for top posts
DM to share your AI project (Due to volume of DMs I'll prioritize subscribers)
579K Followers 53 FollowingThe Gemini app turns research into reality, bringing frontier AI experiences like Omni, Deep Think, Nano Banana, and more to hundreds of millions of people.
116K Followers 17 FollowingImpossible? Let’s see. From algorithms to neuroscience to AI, Google Research strives to progress science, advance society & improve billions of people’s lives.
3K Followers 578 Following姜力炜 @NVIDIA | prev. @uwnlp @stanford @allen_ai 🧊 a curious mind, an independent thinker, and a lifetime adventurer 🌊 views my own
3K Followers 8K FollowingSoftware,Hardware,Astrophysics, Web3, Stocks ,AI ,MLX,Cuda .In perplexity over answer to life universe and Everything . Founder at https://t.co/R6wDNPGrVc
315 Followers 445 FollowingModel as Product @AntLingAGI. Founder of AntOSS. Seasoned Engineer & Strategist. 10+ yrs of eng exp + MBA. Global citizen lived in China, Singapore, Canada, US.