@deepseek_ai V4.1 Flash is now available on @nebiustf
Here’s a quick demo in the Playground, including a speed test and a vision example.
Links and credits in the reply 👇
I was testing DeepSeek V4.1 Flash on @nebiustf , and honestly, the model is pretty damn good.
So I built a tiny real-time webcam app that takes multiple frames from the camera and uses the model to summarize what it sees.
The whole thing was surprisingly smooth, and the results were quite impressive for such a small setup.
Here’s a quick demo of it:
DeepSeek-V4.1-Flash is now live on Nebius Token Factory
The model comes with native image understanding plus an architecture built for faster inference and higher throughput on long, input-heavy workflows.
It combines native vision and a 1M token context with a 552B MoE that
Want to try open models for coding?
Keep your coding agent - Claude Code, Codex, OpenCode, Cline - and just swap the model underneath.
In this short talk:
* open vs closed models
* why the coding harness matters
* what coding tasks cost
* quick setup demo
Watch:
youtube.com/watch?v=gjNcSo…
2 really good open models - GLM-5.3 and DeepSeek-V4-Pro-0813 now on @nebiustf
pushing intelligence while being pretty economical!
viz from Token Factory model visualizer : sujee.github.io/practical-llm-…
Two new models for coding and agent workflows are now live on Nebius Token Factory.
DeepSeek V4-Pro-0813 is the official V4 Pro release, built for coding agents that use tools, reason through complex problems, and work across multiple steps.
GLM-5.3 brings Z. ai’s latest
As I was updating my model indexes, I noticed that @ArtificialAnlys has updated its Intelligence Index.
The new AA Intelligence Index v4.3 introduces several significant changes, including a tougher set of evaluations.
You can read about the changes here:
x.com/ArtificialAnly…
One immediate effect: scores are lower across the board. That doesn't necessarily mean the models got worse. The benchmark suite got harder, so the scale has effectively shifted.
At the top of the latest leaderboard:
- Claude Fable 5.1: 53 = 66 → 53 = ▼ 13
- GPT-6 Astra: 53 = 61 → 53 = ▼ 8
And the top 3 open-weight models:
1. GLM-5.3: 45
2. Kimi K3: 44 (60 → 44 = ▼ 16)
3. GLM-5.3-Flash: 42 (57 → 42 = ▼ 15 )
The score changes for open models are pretty substantial.
See the images below 👇 for the changes and the latest leaderboard.
As they say, it's all relative 😀
What matters most is how models compare against each other under the same evaluation suite.
Announcing Artificial Analysis Intelligence Index v4.3, upgrading Terminal-Bench to 4.0 and adding AutomationBench-AA, an agentic workflow automation benchmark with a private test set. This is a continuation of our rollout of Intelligence Index v5
Changelog (Index v4.2 → Index
@ArtificialAnlys I was just refreshing my open-model rankings and saw the shift in scores
Kimi K3: 60 → 44
GLM-5.3-Flash: 57 → 42
GLM-5.2: 53 → 39
A good reminder that the absolute number matters less than the relative ranking under the same benchmark suite.
Had a fun time geeking out and building with @dN0t and @devopsjacquie
- using open models for coding -- they are surprisingly good (and cheaper)
- using Hermes to find vintage auto parts!
- new models in @nebiustf
Get started by singing up to the Builder Program : lnkd.in/gkUXHpMS
And share with us what you build.
x.com/i/broadcasts/1…
The Nebius AI Builder Program is now available.
AI isn’t just a model you call anymore. It’s a system you build. And builders, not a few closed labs, will decide what it becomes.
The open ecosystem has all the pieces. We want to make it easier to put them together and start building.
The program is free, with $400+ in credits and discounts, working code and cookbooks, office hours with engineers, and a community to build with.
We’re joined by @NVIDIAAI, @LangChain, @huggingface, @cognition, @OpenHandsDev, @tavilyai, @TolokaAI, @composio, @PrimeIntellect, @MiniMax_AI, @Alibaba_Qwen, and more joining soon.
Join with the link in the comments 👇
Here’s a quick explainer on fine-tuning, distillation, and quantization - using LEGO® bricks 🧱
🟢 Fine-tuning: specialize a pretrained model with new data.
🟠 Distillation: train a student model to learn from a teacher.
🟡 Quantization: use lower-precision weights to reduce memory and bandwidth use.
Video below 👇
More detail:
sujee.dev/post/fine-tuni…
What should I explain next with LEGO: LoRA, pruning, or speculative decoding?
Drop your vote in the comments 👇
Went to buy a couple of 4TB portable SSDs and almost fell off my chair looking at the prices 😳
Damn. See screenshots 👇
@demian_ai has been writing about the memory/storage crunch driven by the AI boom.
It is finally sinking in for me :-)
A few of his posts:
- x.com/demian_ai/stat…
- x.com/demian_ai/stat…
- x.com/demian_ai/stat…
$PENG turned scarcity into revenue.
Samsung made memory stretch.
Micron locked down wafers.
Korea broke the wrapper.
this week, the book looked flat but the stack was being repriced ↓
Under a flat headline, the market made a clear choice: pay for what can turn scarcity into
227 Followers 899 Followingi build crms and sites that book customers
try log out for free @ https://t.co/G2VznryHCh
cheap ai tokens @ https://t.co/vO4lvcBCJx
472 Followers 525 FollowingFreerange capitalist and explorer, Author of All Outcomes Are Acceptable, AI Infrastructure investor - Nebius, Micron, SanDisk, Amazon, Meta
1K Followers 5K FollowingSmall Community Driven AI Hackathons. Build during the week & vote on the weekend. 🛜
Solo builds. Capped small. Free forever.
Yard #3 September 21-27
1K Followers 662 FollowingSenior Dev Rel 🥑 I turn AI SDKs into guides, demos, and working code. 5 yrs AI + blockchain infra. Pilot & flight instructor ✈️ 🇮🇹🇺🇸🇧🇷
74K Followers 1K FollowingArmin's handler at https://t.co/B05ybKGkzx. Old man yelling at Claudes.
https://t.co/Q1wG57v1yc
https://t.co/mnOoWUr0TO
https://t.co/8i5vIRE0Wn
8K Followers 2K FollowingWorking on math AI at https://t.co/u95v5xJnFC and telescope software at https://t.co/Yx0Z8UGvEc. firstflagPOISONed. Formerly: Parse cofounder, Facebook, Google
145K Followers 94 FollowingSane + 🌶️ takes in an insane AI world... AI capabilities researcher: co-created RLHF/ChatGPT @ @openai now trying to right the wrong 🤭 (ceo @typesafeai)
2K Followers 3K Followinghelping startups & VCs think about compute @nebiusai | curiosity-maxxing | all things AI, technology, politics | opinions and banter my own 🙃
7K Followers 2K FollowingCo-Founder, CTO-CPO of @SentoraHQ (fmr IntoTheBlock), Co-Founder of @layerlens_ai, @neuralfabric( acq by Cisco) and The Sequence, Teaching at Columbia-Wharton
1K Followers 5K FollowingSmall Community Driven AI Hackathons. Build during the week & vote on the weekend. 🛜
Solo builds. Capped small. Free forever.
Yard #3 September 21-27
294K Followers 852 FollowingCooking fun AI systems & products @databricks. Prev: co-founder & CTO @ Hyperbolic, OctoAI (acquired by @nvidia) Apache TVM, PhD @ University of Washington.
472 Followers 525 FollowingFreerange capitalist and explorer, Author of All Outcomes Are Acceptable, AI Infrastructure investor - Nebius, Micron, SanDisk, Amazon, Meta
7K Followers 3K FollowingRobot maker by profession, tinkerer by nature. Creator of Raspbian. Looking for people with similar quirky outlooks on life. I try to call things as I see them.
2K Followers 470 FollowingPrincipal Developer Advocate at @AWScloud. I tweet stuff - sometimes AWS, sometimes not. Fan of tech, food, coffee, and other stuff. All opinions my own.