People still underestimate how large the data market will become.
Better models do not reduce demand for data. They increase the number of useful things models can learn.
The frontier keeps moving, so the data has to keep getting harder.
Muse Spark 1.3 is rolling out today with frontier performance almost too cheap to meter. This is the biggest jump we've made so far on coding and agentic work. Try it in Muse Code and our API.
Next up 🍉 and Muse Spark open weights releases coming soon.
Legal and finance are two of the highest-impact domains to crack.
3.8 Flash now leads both Harvey’s Legal Agent Benchmark and Vals Finance Agent v2. Huge congrats to GDM ⚡️
Introducing Gemini 3.8 Flash, another jump in Gemini's agentic + coding capabilities, and our 3rd updated Flash model in only 6 weeks...
This model has been a ton of fun to work with, excited to see what you all think!
This is a historic moment for Korean sovereign AI and the open weights ecosystem writ large. We look forward to supporting Motif and the Korean AI ecosystem as they push the frontier of open intelligence.
Motif 3 leads its comparison set on τ³-Banking, and posts 94.7 on τ²-Bench Telecom, 74.9 on Terminal-Bench 2.1, and 76.2 on SWE-bench Verified. It’s also great to see @nvidia's open foundation put to use here - Motif 3's post-training was done via NeMo-RL.
In a world where it’s easy to assume that the AI race is already won, the team at Motif is proof that the field is open, and we’re still at the starting line.
Today, we are releasing Inkling-Small.
Inkling-Small achieves comparable performance to Inkling at a quarter of its size. It features 276B total parameters, 12B active. We are making the full weights available.
thinkingmachines.ai/news/inkling-s…
Fine-tune it on Tinker today, or chat with
Congrats to the @WeAreLegora team on the release of the Legora BAR, a benchmark for agentic legal work built on 5,161 real law firm cases across 28 practice areas!
@AfterQuery helped QA the benchmark with Legora. Proud to work with their team on measuring the frontier of agentic legal work.
Link to the full blog post in the replies.
Big news! Our new MAI-Cyber-1-Flash model combined with MDASH, our multi agent security harness, delivers 96% on the CyberGym benchmark, 12pts above Mythos, at HALF the cost.
Proud of the team. More details in THREAD:
@AfterQuery is hiring SWEs!
In 14 months, AfterQuery has surpassed $100M in revenue run rate and works with all leading AI labs to build the data and systems that models train on.
Our founding eng team joined from firms like Citadel Securities, Palantir, and Meta, and we're hiring more.
Apply in the link below or drop your email.
Refer a successful hire and earn $10K.
To the researchers impacted by the Amazon AGI layoffs: @AfterQuery is hiring post-training researchers.
We work on data solutions that support all frontier labs. If you’ve thought deeply about what improves model capability or behavior, we want to talk. DMs open.
534 Followers 2K FollowingAuthor of Before Consensus Intelligence
Founder & Principal, W-Axis Lab|
Capital Allocator | Former VP of BITMAIN PoW @BITMAINtech |@Draper_u Alumni
21K Followers 2K FollowingPartner at Avenir, where I invest in startups.
Here to bring some analytical irreverence, while trying to add a little information to the world.
Views my own.
725 Followers 2K FollowingFounder|Content Creator|Story Teller
“Kofa Muse with all your Entertainment Tech News , and Views Tech Bruh 😎 not Tech bro 🤓
download my app link in bio
65K Followers 31K FollowingCopies that attract clients + follow-up methods that close.
Subscribe for the fiction and sales frameworks.
DM = Business = Sub First. No casual chats.
Satire.