Ai-research
16%23 posts analyzed over the last 12 weeks
Over the last 12 weeks
Average across all posts
This month vs. previous
23 posts analyzed over the last 12 weeks
Over the last 12 weeks
Average across all posts
This month vs. previous
Growth slope (6 weeks)
Best week: 8 juin (195 avg. likes)
Maxime Labonne
Head of Post-Training @ Liquid AI
๐ณ Revisiting DeepSeek-V4: how RL got replaced with on-policy distillation DeepSeek released V4 in late April with a huge tech report! I finally found time to finish it and finalize my article on Substack. There are thrโฆ
Maxime Labonne
Head of Post-Training @ Liquid AI
โ๏ธ Multi-teacher on-policy distillation (MOPD) Xiaomi's MiMo team presents a clean recipe for combining several specialist models into one, without the usual cross-domain interference. Instead of pooling all the traininโฆ
Clem Delangue ๐ค
Co-founder & CEO at Hugging Face
Are scaling laws finally working for time series foundation models? Today, Datadog is releasing Toto 2.0 weights in Apache 2.0 on Hugging Face. It's a family of open-weights TSFMs from 4M to 2.5B parameters, where everyโฆ
Sandipan Bhaumik
Data & AI Technical Lead | Production AI for Regulated Industries | Founder, AgentBuild
Claude Code Isn't a Chatbot. ๐๐ญ'๐ฌ ๐ ๐๐ ๐๐ง๐ญ๐ฌ ๐๐ฎ๐ง๐ง๐ข๐ง๐ ๐๐ฌ ๐ ๐๐ข๐ฉ๐๐ฅ๐ข๐ง๐. Most engineers using Claude Code don't realize this. They think they're prompting one model. What's actually happening: A โฆ
Kevin Degila
Head of Data and AI | Forbes 30 under 30
Vous avez รฉtรฉ nombreux ร precommander le livre aprรจs son annonce hier. Merci pour cet accueil enthousiaste. ยซ Construire un LLM de zรฉro - des maths aux agents IA ยป, c'est 4 parties, 21 chapitres, et une promesse : partiโฆ
Ghita Houir Alami
Co-Founder and CEO at ZeroEntropy (YC W25) - We're Hiring!
Mini models are supposed to get cheaper every release. They're getting more expensive. Gemini 1.5 Flash to 2.5 Flash: 4x on input, 8x+ on output. Gemini 3 Flash to 3.5 Flash: another 3x. And 3.5 Flash, the model that'sโฆ
Ghita Houir Alami
Co-Founder and CEO at ZeroEntropy (YC W25) - We're Hiring!
MTEB largely utilizes binary relevance scores: a document is either relevant or it isn't. But under binary labels, metrics like NDCG degenerate. It can't tell a perfect answer apart from a loosely related document. Soโฆ
Ben Burtenshaw
Community Education in AI @ Hugging Face
Had loads of fun setting up continuous learning through on agent traces. This post is based on a script that implements continuous learning with self distillation on agent traces. It's purely an experiment and not sometโฆ
Pini Reznik
CEO and Co-Founder @ re:cinq | AI Native Transformations
Last week we ran the 9th edition of AI Native Netherlands at Adyen, with more than 100 people in the room. Many thanks to Kwok He Chu and Ayodeji Ogundare at Adyen to help organise this event. ๐ช๐ฒ ๐ต๐ฎ๐ฑ ๐๐๐ผ ๐ด๐ฟ๐ฒโฆ
Max Buckley
Founding something new
Exa at AI Tinkerers Dublin on Thursday May 28th ๐ฎ๐ช I'll be representing Exa in my home country of Ireland very soon. I'll be presenting at AI Tinkerers Dublin, hosted by Baseline, giving a deep dive on building AI resโฆ