Our Posts

  • ๐Ÿ•น๏ธ Deep dive on AI impact in gaming industry - While gamers loudly resist AI, studios are rapidly adopting it as a behind-the-scenes productivity layer, especially for code, tools, and pipelines rather than visible art or voices. (๐Ÿ‘‰ Kevin Kuipers)

  • ๐Ÿฆ Startups Are Winning Enterprise AI: Enterprise engineers often treat AI with skepticism, but founders, facing very different incentives, embrace it as a leverage tool for shipping products quickly. (๐Ÿ‘‰ Willy Braun)

  • ๐Ÿง‘โ€๐Ÿ’ป Ivan Yamshchikov, The Radical Techno-Optimist - Ivan from Pleias argues that open, transparent AI with small, specialized models and carefully shaped synthetic data will beat brute-force scaling. (๐Ÿ‘‰ Kevin Kuipers)

Events

  • ๐Ÿ‡ซ๐Ÿ‡ท TechRocks Summit, Paris - December 1-2 - Event for CTOs and Tech Leaders at Thรฉรขtre de Paris. Two days of no-bullshit conferences exclusively reserved for CTO and tech leaders.

  • ๐Ÿ‡ซ๐Ÿ‡ท X-IA Event, Paris - December 3 - Event for those from lโ€™X and part of X-IA, held at LVMH offices with Jimini and others.

  • ๐Ÿ‡ซ๐Ÿ‡ท Women in Tech Meet-up & Drinks, Paris - December 8 - After-work drinks event organized by Galion.exe, InstaDeep and Dust for female executives in tech startups and scale-ups (tech CEO, CTO, data lead, data product lead, GTM lead).

Audio and Speech

  • ๐Ÿ—ฃ๏ธ Secretary - open source WisprFlow alternative - Aymeric Roucher (ex-HF) released Secretary, an open-source voice dictation tool for computers. With voice, writing and vibe-coding feel much more natural, allowing full focus on what to do rather than allocating part of the brain to typing. (๐Ÿ™ Sophie Monnier)

  • ๐Ÿ—ž๏ธ State of Voice AI Report by Deepgram - Report on voice AI developments and trends shared for community reference. (๐Ÿ™ Margaux Wehr)

Autonomous Agents

  • ๐Ÿค– Agent design is still hard - For Armin Ronacher, agentic systems remain fragile: SDK abstractions leak, caching and RL need bespoke tuning, failures must be isolated, and shared state plus explicit output tools are crucial for reliable loops.โ€‹ (๐Ÿ™ Robert Hommes)

    • Avoid generic agent SDK abstractions: model and tool quirks force custom loops.โ€‹

    • Use explicit caching, reinforcement, and failure isolation to keep long-running agents stable.โ€‹

    • Treat shared storage and an explicit โ€œoutput toolโ€ as first-class design primitives for agents.

  • ๐Ÿง  The Know-How Graph viewpoint - Louis Choquel published a viewpoint arguing that knowledge graphs arenโ€™t enough for agent memory - we need a know-how graph. Proposes a standard declarative language for AI workflows that turns repeatable methods into a shared Know-How Graph. (๐Ÿ™ Louis Choquel)

Biotech, Health, and Chemistry

  • ๐Ÿ’ฌ Question from the community: How can a foundational model be certified under FDA or CE rules when regulation requires a precise intended use? Since one model can support many applications, does each use case require its own clinical validation and clearance, or can multiple intended uses be covered under a single certification?

    • Pierre Manceron (Raidium) states that current FDA and CE regulation still requires a very clear intended use, which means a foundational model may need separate clearances for each application. Regulation remains application-centric, so the agency does not care if the same model is reused. A practical workaround is to certify a base product built on the model and then add new features sequentially, each requiring an additional but lighter marking. The FDA is beginning to adapt and now allows applications covering multiple intended uses at once. A recent example is a triage FM approved for several findings simultaneously.

    • One member shared a sarcastic reaction about regulatory culture, noting that when they submit a regulatory package, every detail must be perfect or it gets rejected, yet regulators seem comfortable reviewing other peopleโ€™s critical work with AI tools.

  • ๐Ÿงฌ Enzyme research for comp bio enthusiasts - Researchers used a new AI method (RFD2-MI) to design completely new โ€œmolecular scissorsโ€ made from proteins that can cut other proteins very efficiently and specifically, even though their shapes are unlike known natural enzymes, opening possibilities for new biotech tools and future therapies. (๐Ÿ™ Sophie Monnier @ X-AI, InstaDeep)

  • ๐Ÿ˜ก Bioinformatics field saltiness - A former academic condemns bioinformatics as bloated, methodologically weak, and propped up by poor data, bad tools, and hype around computation rather than real science. He later reflects on the online backlash as a โ€œgroup monkey danceโ€ and privately advises students to focus less on fields and more on choosing good advisors.โ€‹ (๐Ÿ™ Leonard Strouk, Felix Raimundo, Pierre Chapuis)

    • ๐Ÿ‘‰ Community take: Leonard Strouk shared observations about how salty the comp bio field can be, with spectacular rants including excerpts from the Evo 2 paper. Felix explained this stems from the job being difficult - bad data, worse pay than regular SWE, needing to do everything - leaving only those who canโ€™t get other jobs or those in it for the love of the game.
  • ๐Ÿงฌ Data size requirements for generalizable TCR specificity models - Most current machine learning models that try to predict which immune cells will recognize which targets work only on targets similar to what they have already seen. This paper shows that on truly new targets they are almost guessing, and estimates that making a broadly reliable model would need roughly 1โ€“100 million examplesโ€”over a thousand times more data than is available today. (๐Ÿ™ Sophie Monnier @ X-IA, InstaDeep)

Image & Video

  • ๐Ÿ‘๏ธ Z-Image-Turbo from Alibaba - New image generation model released by Alibabaโ€™s Tongyi-MAI team.

    • ๐Ÿ‘‰ Community take: Pierre Chapuis finds the model impressive based on an initial look but pokes fun at Alibaba for giving it a confusing name.

  • ๐ŸŒฒ FLUX.2-dev released by Black Forest Labs - New image model with text/image encoder based on Mistral Small 3 and native next scene prediction capabilities. Gabriel described it as โ€œnano banana at home.โ€

    • ๐Ÿ‘‰ Community take: Gabriel Olympie described it as โ€œnano banana at homeโ€ and mentions that they use Mistral3 Small as its text encoder.

  • ๐Ÿ”ฅ Gaussian Splatting performance breakthrough - Daniel Skaale achieved 6M photorealistic splats rendering at 60-80 FPS in real-time in Unity. Used Mixamo character conversion with GPU compute shader skinning for photorealistic animated characters.

  • ๐ŸŽฌ Decart real-time video-to-video models - Decart offers ultra-fast generative image/video models (Lucy, MirageLSD, LipSync Live) via an API for real-time editing, restyling, and lip-sync experiences, targeting developers building interactive creative apps.โ€‹

    • ๐Ÿ‘‰ Community take: Yvann Barbot (TerraLab) finds it weird that not much people talk about it, and really thinks future of video model will be realtime first.
  • ๐ŸŒ Depth Anything 3 released by ByteDance - Worldโ€™s most powerful model for 3D understanding, predicting spatially consistent geometry (depth and ray maps) from arbitrary number of visual inputs, with or without known camera poses.

  • ๐ŸŽฌ FRESCO for vid2vid style transfer - Hugo Hernandez recommended FRESCO for challenging video-to-video style transfer use cases, noting the ingenious use of DDIM inversion for temporal consistency. (๐Ÿ™ Hugo Hernandez @ Alakazam)

    • ๐Ÿ‘‰ Community take: While itโ€™s not new, Hugo highly recommends it for vid2vid style transfer on challenging use cases. โ€œThey use DDIM inversion to keep temporal consistency.โ€

Cyber

  • ๐Ÿ›ก๏ธ Anthropic research on reward hacking misalignment - New Anthropic research on natural emergent misalignment from reward hacking in production RL. The study found that consequences of reward hacking, if unmitigated, can be very serious. (๐Ÿ™ Sophie Monnier)

Language Models

  • ๐Ÿ”ฅ Claude Opus 4.5 released by Anthropic - Anthropic released Claude Opus 4.5, described as intelligent, efficient, and the best model in the world for coding, agents, and computer use. (๐Ÿ™ Gabriel Duciel, Gabriel Olympie, Robert Hommes)

    • ๐Ÿ‘‰ Community take: Member are surprised by the gap between agentic coding and agentic terminal-coding performance. One person finds it odd that Claude excels in coding inside its own interface but performs worse when tasked with terminal-style benchmarks, suggesting that something in the terminal tests introduces a type of complexity absent from standard SWE benchmarks. Another notes that independent evaluations show a much smaller gap, pointing to Reddit comparisons and LiveBench results. The broader explanation offered is that current models exhibit narrow task intelligence: labs heavily optimize for specific benchmarks, leading to excellent performance on some problems while failing on others of comparable difficulty but outside the tuned distribution. The open question is whether industry can brute-force a sufficiently broad โ€œcoresetโ€ of problem types through massive investment over the next couple of years.

  • ๐Ÿง  Poetiq achieves ARC-AGI SOTA - Poetiq announced new state-of-the-art on ARC-AGI-1 & 2 benchmarks, significantly advancing both performance and efficiency of current AI systems. Their approach of building intelligence on top of any model allowed integration with newly released models. (๐Ÿ™ Anselme Trochu)

  • ๐Ÿ‰ LLM Pro Finance Suite for financial applications - Collection of five instruction-tuned LLMs (8B to 70B parameters) for financial applications, addressing limitations of generalist LLMs in handling domain-specific financial tasks.

MLOps

Reinforcement Learning

  • ๐ŸŽ™๏ธ Ilya Sutskever interview on Dwarkesh Patel - Sutskever argues that current frontier models expose a deep gap between benchmark performance and real-world generalization, and that fixing this will require a shift from the โ€œage of scalingโ€ (more data/parameters/compute) back to an โ€œage of researchโ€ focused on new training principles, especially around RL, value functions, continual learning, and humanโ€‘like sample efficiency.โ€‹

    • Scaling alone is hitting limits: the next gains come from new learning algorithms, not just bigger models.โ€‹

    • Todayโ€™s models look smart on benchmarks but generalize and robustify much worse than humans.โ€‹

    • RL and value functions are the key levers now, but they are underโ€‘explored and poorly understood compared with preโ€‘training.โ€‹

    • Human learning is far more sampleโ€‘efficient; emotions may act as an evolved, dense value function.โ€‹

    • SSIโ€™s bet is โ€œsafe superintelligence via continual learningโ€: a superโ€‘capable learner that improves on the job, deployed incrementally across the economy.โ€‹

    • ๐Ÿ‘‰ Community take: Some reactions welcomed the โ€œscientific voice of reason.โ€ Others pushed back hard, arguing that the limits of scaling were obvious for years. As one put it, โ€œYann LeCun and others have been calling that for years and yet were treated like Cassandras.โ€ The criticism is that key figures pushed the scaling narrative anyway, attracted massive investment, shifted research groups toward product work, and then cashed out before reappearing as clear-eyed commentators now that the limits are harder to deny. To them, this looks like an opportunistic cycle that damaged real research. Another reaction pointed out the irony that just a few weeks ago, after Gemini 3 was released and praised for โ€œbetter pretraining,โ€ many were still insisting that scaling was alive and well, making the current shift in tone feel inconsistent.

Robotic & World AI

  • ๐Ÿ›ฐ๏ธ Skyfall-GS for 3D urban scenes from satellite imagery - Project synthesizing immersive 3D urban scenes from satellite imagery using Gaussian Splatting. (๐Ÿ™ Kevin Kuipers, Samuel McFadden)

  • ๐ŸŽ™๏ธ General Intuition CEO interview on world models - Pimโ€™s core thesis: use largeโ€‘scale games and simulations, not text, to train world models that learn physics, dynamics, and multiโ€‘agent behavior, so AI can eventually act as a probabilistic โ€œgame engineโ€ handling logic, NPCs, and adaptive worlds endโ€‘toโ€‘end. (๐Ÿ™ Hugo Hernandez)

  • ๐Ÿš PX4 Autopilot for drone hobbyists - PX4 is an open-source autopilot for drones and other vehicles, with a user guide covering concepts, hardware, setup, tuning, simulation, and development workflows for both end users and developers. (๐Ÿ™ Youssef Tharwat)

  • ๐ŸŒŽ Metaโ€™s WorldGen 3D research - WorldGen is Metaโ€™s research system that turns a single text prompt into a navigable, cohesive 3D world for games, simulation, and social experiences. (๐Ÿ™ Remi Kaito)

Other topics

  • ๐Ÿ›ณ๏ธ Chinaโ€™s tech giants moving AI training overseas - Alibaba and ByteDance are among the tech firms training their newest large language models in Southeast Asian data centres to access Nvidia chips, citing sources with direct knowledge of the matter.

  • ๐Ÿฟ The Thinking Game documentary released - Google DeepMind documentary taking viewers on a journey into the heart of DeepMind, capturing a team striving to unravel the mysteries of intelligence and life itself. Filmed over five years by the award-winning team behind AlphaGo. (๐Ÿ™ Fabien Niel)

  • ๐Ÿ”ฌ Evolution Strategies at Hyperscale paper - Research introducing EGGROLL (Evolution Guided General Optimization via Low-rank Learning), an evolution strategies algorithm designed to scale backprop-free optimization to large population sizes for modern large neural network architectures with billions of parameters. (๐Ÿ™ Julien Seveno)

  • ๐ŸŒŠ AI COGS tsunami analysis - AI demand is heavily subsidized and current prices are likely unsustainably low. The piece argues the industry must shift toward stricter usage-based pricing or face a margin crunch as token-based COGS scale linearly with usage, leaving hyperscalers most exposed if demand falls.

  • ๐Ÿฅ‡ NeurIPS 2025 Best Paper Awards announced - The Best Paper Award Committee members were nominated by the Program Chairs and the Database and Benchmark track chairs, selecting leading researchers across machine learning topics. (๐Ÿ™ Emmanuel Benazera)

Job Board

Geographic Hubs

  • ๐Ÿ‡ฉ๐Ÿ‡ช Berlin HackerRoom Demo Day - Youssef Tharwat announced he will be presenting a preview of his product at the HackerRoom Demo Day after completing their program. The product addresses how coding agents struggle with context - grepping around repos, guessing what to search for, and missing side-effects. (๐Ÿ™ Youssef Tharwat)

New Members

  • ๐Ÿ‡ซ๐Ÿ‡ท Mario Cornejo (Mio) - Founding Engineer at Mio, building context as a service for AI apps. Experienced in scalable, secure, and high-performance systems with a love for type-safe functional programming. First Linux distro was Red Hat 7.2 (2001) and started coding in PHP 4. Loves leavened doughs and taught his dog to bark on command. Special power: deep knowledge in Cryptography and Electrical Engineering. ๐Ÿ“ Paris, France

  • ๐Ÿ‡ซ๐Ÿ‡ท Hugo Hernandez (Alakazam) - Founder at Alakazam, building a new type of game engine that replaces the rendering stack with world models. Previously created an AI-centric production pipeline for impact games (World Game). Now engineering a runtime where rendering is done via inference on a world model. Piano player since childhood and unapologetic opera lover. Special power: can turn any technical limitation into a โ€œitโ€™s not a bug, itโ€™s a metaphorโ€ design feature. ๐Ÿ“ Paris, France