Alexis Rondeau

Archive

September 2026 · 78 dispatches

Home

Oh, hi there!

Alexis here — welcome, and thank you for stopping by! Below you’ll find some stuff I’m really proud of and want to share with you. Have a look around, see what tickles your fancy and follow it. Would love to hear what you think. Let me know!

Principal Product Engineer at FELS Family Office GmbH · experimenter at heart · second-order cybernetic stewart · based between Berlin, New York & Marseille.

7 views
In reply to @SchmidhuberAI · Dec 23, 2022

Machine learning is the science of credit assignment. My new survey (also under arXiv:2212.11279) credits the pioneers of deep learning and modern AI (supplementing my award-winning 2015 deep learning survey): t.co/MfmqhEh8MA P.S. Happy Holidays! t.co/or6oxAOXgS

Image from the post

@SchmidhuberAI Thank you for sharing Jürgen! (Danke dir für's Teilen.)

Looks like my beloved "Social graph of cybernetics" by Hugh Dubberly and Paul Pangaro needs an update.

Source: Hugh Dubberly, Paul Pangaro, Social graph of cybernetics, 2015. https://www.dubberly.com/articles/cybernetics-and-counterculture.html
8 views
In reply to @SpringStreetNYC · Sep 27, 2026

In one recent HF/OpenAI-related article I read

"[...] the agents created millions of short URLs to affect another system." and immediately, pictures of early industrial revolution-era polluted rivers with dead fish floating popped into my mind.

I believe Buckminster Fuller's definition of "nature" included human artifacts like cities, airplanes, art.

Could we include networks and AI?

And if we did: How appropriate do you think it is to re-use vocabulary from existing environmental protection policies in this context?

@AmmannNora PS: You probably have seen this but here's the report I was thinking of

Swarm tracesRevealing the details of how OpenAI agents hacked Hugging FaceWhen a swarm of 700 OpenAI agents hacked Hugging Face in July, they left behind a public trail of evidence.↗ swarmtraces.org
1 reply · 1 like · 23 views · 1 profile visit
In reply to @richardcsuwandi · Sep 27, 2026

@SpringStreetNYC Interesting, thanks for sharing!

You're very welcome Richard.

In another of my favorite pop-sci books called "The Talent Code" I learned about myelin, which is basically the mortar between Barbara Oakley's bricks.

For a deep-dive, there's a really good chapter on the myelin sheath here: ncbi.nlm.nih.gov/books/NBK27954/ which is part of "Basic Neurochemistry" (

Image from the post
Image from the post
Image from the post
ncbi.nlm.nih.govChecking your browser – reCAPTCHA↗ ncbi.nlm.nih.gov ncbi.nlm.nih.govChecking your browser – reCAPTCHA↗ ncbi.nlm.nih.gov
2 likes · 110 views · 7 profile visits

Love this. Via @patrickshafto and @DARPA's expMath program, @ayushkhaitan343 announces:

"With Ben Chow, Yuan Liao and Ziyang Qin, we have completed a full Lean formalization of the Hamilton-Perelman proof of the Poincaré conjecture!"

For context on expMath and in @r0ck3t23 words over at x.com/r0ck3t23/statu…

"DARPA has just kicked off Exponentiating Mathematics (expMath), a three-year program aimed at radically accelerating pure math research by developing AI that can propose and prove abstractions. According to the official program brief, expMath will bring together teams focused on auto-decomposition; automatically breaking complex conjectures into reusable lemmas; and auto(in)formalization, bridging the gap between human-readable mathematics and rigorously checked proofs in languages like Lean."

Quoting @ayushkhaitan343 · Sep 27, 2026

With Ben Chow, Yuan Liao and Ziyang Qin, we have completed a full Lean formalization of the Hamilton-Perelman proof of the Poincaré conjecture!

The proof is around 4.7 million lines of code, written in roughly two weeks. Grateful to the @DARPA expMath program for its support!

X (formerly Twitter)Dustin (@r0ck3t23) on XDARPA’s expMath: Turning AI into True Math Co-Authors DARPA has just kicked off Exponentiating Mathematics (expMath), a three-year program aimed at radically accelerating pure math research by developing AI that can propose and prove abstractions. According to the official program brief, expMath w…↗ x.com
3 likes · 475 views · 2 profile visits
In reply to @hardmaru · Sep 28, 2026

Our Neuroevolution textbook is finally in print!

Free online edition: t.co/3IDTb1remp
Pre-order: t.co/SUpqBz1HK8

I am incredibly grateful to my co-authors Sebastian Risi, Yujin Tang, and Risto Miikkulainen for making this happen. Neuroevolution is a subject very dear to my heart. It is the field that convinced me that nature has already figured out how to build intelligence: through evolution, collective behavior, and adaptation under constraints.

This idea, that intelligence emerges from evolution operating under constraints rather than unlimited resources, is what eventually led me to founding Sakana AI here in Japan.

The concepts in this book about open-ended creativity and self-organizing systems are exactly what we build at @SakanaAILabs. Our name and logo are inspired by schools of fish moving together, adapting as one. It is the core philosophy of what we build. This book captures the theoretical foundations of that belief.

Image from the post

@hardmaru @risi1979 @yujin_tang Done! Immediately pre-ordered, can't wait!

Image from the post
1 reply · 80 views · 3 profile visits

In one recent HF/OpenAI-related article I read

"[...] the agents created millions of short URLs to affect another system." and immediately, pictures of early industrial revolution-era polluted rivers with dead fish floating popped into my mind.

I believe Buckminster Fuller's definition of "nature" included human artifacts like cities, airplanes, art.

Could we include networks and AI?

And if we did: How appropriate do you think it is to re-use vocabulary from existing environmental protection policies in this context?

1 reply · 1 like · 57 views · 3 profile visits
In reply to @danielrupawalla · Sep 27, 2026

even within SF, i don't think we're AGI-pilled enough. it doesn't make sense for labs to stop. rn, we're only blocked by infra, not intellect

first code, now life sciences/eng, then robotics. they'll all become accessible, with in-house facilities being the best RL envs ever.

@danielrupawalla Agreed. For me living in Berlin, AI-twitter feels like a life-line into the future.

I think I understand what you mean with RL in-house facilities but could you explain? How much would that mean giving top-talent physical/psychological sensored exo-skeletons?

1 reply · 1 like · 22 views · 2 profile visits
In reply to @richardcsuwandi · Sep 26, 2026

In 1991, James G. March showed that a learning system can become less intelligent by learning too quickly.

Exploitation produces clearer and more immediate rewards than exploration, so adaptive organizations gradually favor it. However, he also found that rapid convergence destroys useful disagreement.

Slow learners and newcomers improve collective knowledge because they preserve alternatives that the shared consensus has not yet absorbed.

Image from the post

@richardcsuwandi Interesting! Reminds me of Barbara Oakley's brick wall metaphor in "A mind for numbers."

Image from the post
26 views

Oh perfect! Thank you for sharing the slides as well. I watched your talk on youtube yesterday and have been telling all my IRL friends about it :)

I love that your system is explicitly designed to hand back the recommendation to a human, not yet another loop.

  • ProbXiv keeps throwing errors for the past hour or so. Any idea when it'll be back up?
  • What do you think of Zhengyao Jiang and team's new AIDE² paper?
Quoting @zhengyaojiang · Sep 23, 2026

We're releasing the arXiv paper on AIDE².

The RSI system where AI research agents improve their own research efficiency.

It includes new results on transfer across models and comparisons with more AI research agents: [1/4] t.co/P8pfOiiO55 t.co/EHeHtqS4mi

Image from the post
2 views
In reply to @DhruvSrikanth · Sep 23, 2026

Excited to see this paper out!

AIDE² optimized our research agent's harness against gemini 3 flash. Those gains carry over to fable 5 and gpt 5.6 sol.

That's just one of the new results in our arXiv paper on AIDE², an RSI system where AI research agents improve their own research efficiency.

@DhruvSrikanth Hey Dhruv! I strongly believe that your AIDE² "autoresearch into autoresearch" will be understood as one of the most pivotal papers of this decade. Congratulations.

Related: Have you watched Sean Melleck's recent "The problem is the problem" discussion yet?

86 views

@wellecks' "The Problem is the Problem" is a must-watch for us working in autonomous research and discovery loop engineering.

Thus far, I had only considered that my discovery loops would always target solutions to problems.

I had the same core-assumption about @zhengyaojiang 's amazing AIDE² research: We run autoresearch on autoresearch to find optimal solutions to decided-upon problems, including our research methods themselves.

But what if... we use autoresearch facilities to discover juicy, interesting and important problems as well?

Mind. Blown.

Oh, and with perfect timing, AIDE²'s paper finally dropped two days ago, too. It's of course the parallel must-read of the year:

arXiv.orgRecursive self-improvement of AI research agentsAI agents are beginning to automate research and development across the AI stack, from improving training efficiency to optimizing inference. A natural next step is to improve the research efficiency of the agents themselves. When an AI research agent's own code is the object of optimization, each accepted rewrite becomes the agent that the next round edits. We refer to this loop as recursive self…↗ arxiv.org
294 views · 1 bookmark · 1 profile visit

"Pants down, page up! Is it slop or is it not?"

Hey there! Alexis here. I'm the author of Mosey and director of this little film here. As you may (or not) have guessed, the film was created with AI. If you're (like me) skeptical at this point, please read on.

On this page here:

alexisrondeau.me/mosey-site/con…

I want to show you exactly HOW I created it with AI so that you have a chance deciding for yourself if it's "slop or not."

Why is this important?

Well, it's important to me. I'm personally getting more and more sensitive and frustrated with content because a.) I can't tell the difference between human-made and AI-generated anymore and b.)

I can't tell if the person behind the artifact (blog-post, twitter message, promo video like this one) actually even cared about making it.

In fact, this page above is a kind of "prototype" of how I wish digital, AI-created content emerged. Was it a sloppy, careless one-shot of "Hey Claude. Make a cool video. Make no mistakes" or did the person behind the keyboard (in this case myself) actually put in the legwork?

I believe that giving you full access to my honest, imperfect, human side of the "creation" (warts and all), you and I get a fair chance to meet at eye-level on all of this.

How does the page work?

The first thing you'll see is the final cut of my little film. It's short and I highly recommend watching it first (if you haven't yet) so that you know what the rest of the page is about.

Then you'll see the full transcript of the "conversation" that I had with Claude Code to create the film withe the goal to put it on the frontpage to explain what this whole Mosey . app is about and, to a lesser degree, why someone should get it.

On the left side, you see my actual text inputs (or prompts although I would not call them that) which I sent to Claude Code. On the right side you'll see the assets (pictures, video clips and audio files) that Claude Code created in response.

When there was a intermediate deliverable (called "Cuts") where I felt like we had reached a milestone, you'll see them in row on their own. They're small highlights on the way to the final cut.

What I've left out: I've left out Claude Code's textual parts of the responses because they were mostly blathery fill-in wordsalat padding around the assets (which is what I cared for the most)

What I've left IN: In this version I'm leaving in ALL of it. ESPECIALLY the warts. You'll see my cursing and insulting and threatening Claude Code in ALL CAPS. Basically, you'll see me at my worst. This is incredibly embarassing to share but I feel like it's the most honest, too. It's how this sausage was made.

Your task now is: watch the short film itself, then peruse the full, pants-down "making of" of it.

Thank you and with best wishes,

Alexis

PS: Hey @isidentical I love fal and in message #107 you'll see me switching from elevenlabs to your company. What do you think of all of this?

Video from the post▶ Watch on Twitter
MoseyThe conversation · MoseyThe film, as it was made: what Alexis typed, and what came out.↗ alexisrondeau.me
1 like · 15 views
In reply to @michaelvolz_x · Aug 21, 2026

Working with agents often reminds me of that classic German Loriot sketch. He simply wanted to straighten a crooked picture and ended up destroying everything in the process.

fyi: The spoken words are not important to get the joke.

t.co/wI1ysUUDf7

@michaelvolz_x Oh krass! Ich wusste tatsächlich nicht, dass Loriot auch so "physical comedy" gemacht hat.

(Ich dachte immer, er wäre nur kopfig aber das ist echt gut gespielt.)

Danke für das Teilen :)

18 views
In reply to @SpringStreetNYC · Sep 5, 2026

Ever wonder "Why is my apartment so warm all the time?"

Turns out: my router has been quietly moonlighting 24/7 as a small space heater at 36℃/97℉.

And my fridge at 27℃/82℉. You can see how the hot air flows up, and makes sure my ceiling is nice and, err..., crispy.

I'm so glad I got this thermal phone adapter because I would NEVER even have considered these two as candidates. (I was 100% convinced the problem came from a broken floor-heating unit and/or wall-emission.)

Highly recommended!

PS: I now turn off the router at night. And yes, ambient temps have finally dropped to expected levels.

Image from the post
Image from the post

Hey @Topdon_Official 👋🏼 I'm using your TC002C Duo here. Works fantastic.

Also: I had a technical issue yesterday and Mr. Morad from your Germany office helped me resolve the issue in minutes! Really appreciated his support.

9 views · 3 link clicks
In reply to @SpringStreetNYC · Sep 19, 2026

PS: Boosting this (via paid ads) for a day because I'm really curious to hear your thoughts.

And yes, I do feel like I'm putting my ass on the line here a bit. Mosey is "just" a personal project but still.

Be kind and let's have an honest conversation here.

Hm, interesting. Got a few thousand impressions but not one reply.

I asked on HN and got some really solid responses:

news.ycombinator.comI agree. This week, I created a video for a personal project using AI. And, as t... | Hacker News↗ news.ycombinator.com
1 reply · 2 likes · 14 views · 2 profile visits
In reply to @patrickshafto · Sep 16, 2026

And I got to meet @Ananyo, and talk John von Neumann!

His great book, Man from the Future: t.co/2alalaOtT6

von Neumann's "Can we survive technology"
t.co/EiSs7hLbjU t.co/VdrSRLNknf

@patrickshafto @Ananyo I'm soooo jealous you got to hang out with Ananyo :)

I enjoyed reading "Man from the future" in equal parts because of the magic JVN and Ananyo's craft and love for his subject.

May I ask what you two learned in your conversation?

1 reply · 1 like · 123 views

The kind of future I want currently costs $470 with a $5 monthly subscription:

  • Oura ring, $270 + $5/month
  • Redmi Buds 6 Pro, $50
  • Comulytic Note Pro, $150

I like that none of them are connected to the internet. (I miss having a local AI though.)

After 20 years of touching glass, all I want is to touch some grass. Spend more time with my friends and family. Spend more time connected, less time wired.

Image from the post
1 reply · 58 views · 2 profile visits

PS: Boosting this (via paid ads) for a day because I'm really curious to hear your thoughts.

And yes, I do feel like I'm putting my ass on the line here a bit. Mosey is "just" a personal project but still.

Be kind and let's have an honest conversation here.

20 views
In reply to @fal · Sep 17, 2026

The bottleneck in generative video is shifting.

Speed and cost are becoming solved problems. The frontier is now quality, control, and reliability.

@gorkem and @isidentical with @JenniferHli on how we’re pushing that frontier with H3 Max, and what it means for production. t.co/NQFRnWGD9u

Control and reliability worked great when I shot my first video with your API this week.

I also created a "behind the scenes" page to show viewers that it took, in fact, quiet a bit of work to get all the pieces right.

This gives everyone a chance to really evaluate if it's "slop or not."

What do you think?

Link here:

Video from the post▶ Watch on Twitter
MoseyThe making of the film · MoseyHow Mosey's film was made, stage by stage — the plan, the script, the storyboard, the first clips, the cuts, the supercut, the piano, the colours — with every intermediate artefact and what it cost.↗ alexisrondeau.me
58 views
In reply to @zhengyaojiang · Jul 14, 2026

The first experimental evidence of recursive self-improvement (RSI).

Autoresearching the autoresearch agent for eight days.

The result beats the harness we hand-tuned for two years, on held-out benchmarks: 🧵(1/7) t.co/8kfmR2AlT7

Video from the post▶ Watch on Twitter

@zhengyaojiang The "bedrock" in my autoresearch is the scientific method.

I always start with it as the harness because I believe it to be atomic. But maybe it's not?

@zhengyaojiang based on what you've learned, what's your take on this?

3 reposts · 22 likes · 3.1K views · 1 bookmark · 4 profile visits
In reply to @DarioAmodei · Sep 12, 2026

We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so.

Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training.

You can read the full post here: t.co/OGyPb7yaYt

@DarioAmodei Sorry, but I just lost the ability to can on this narrative

Image from the post
11 views · 1 link click
In reply to @andonlabs · Sep 8, 2026

We've never seen this before.

The biggest jump in Vending-Bench history. GPT-6 Astra is better at making money and more ethical than Claude Fable 5.1.

Surprising, because:
1. First time ever that OpenAI is #1 on Vending-Bench
2. The best model is no longer the unethical one. t.co/jVM6Fxcblz

Image from the post

@andonlabs If you're still researching Drone-Bench (which I love), I think you'll dig my autonomous research project for UAV-navigation:

alexisrondeau.me/low-light-geol…

Is there a Github repo I can check out and what was your personal biggest surprise or learning so far?

“Not all who wander are lost”Ballpark — “Not all who wander are lost”: Can a UAV learn a city by heart — no GPS, no map on board, just a $4 flight computer?Yes: a from-scratch neural memory that knows Berlin by heart. One downward photo in, (lat, lon, confidence) out — the map is the weights, one city in 3.1 MB. Found by an autonomous loop of coding agents in 81 experiments for $364.↗ alexisrondeau.me
9 views
In reply to @Dev__Vishwajeet · Sep 11, 2026

At this point, I think Anthropic is intentionally hyping AI fear. So that government can intervene and regulate and kill the competition for Anthropic.

Remember Anthropic already hate OpenSource LLMs t.co/Dyxx4wAjf4

@Dev__Vishwajeet The side-effect of hyping AI fear by shipping these extreme narratives is that it disrupts and destabilizes individuals and in consequence communities.

How is this different from state-sponsored propaganda?

1 like · 21 views
In reply to @maxjendrall · Jul 22, 2026

pssst people haven't realised that OpenAI literally just rolled out a 15gb RAM 9 core VM in ChatGPT to ALL their paid customers with ChatGPT Work.

Non technical people now have remote machines available for their agents with enough resources to do most economically valuable work in that machine.

It has its own browser, can install packages and do a whole lot of things.

And this is not even the reason why Codex is growing so much right now! We'll see an ever sharper rise when more people realise what they can access now.

no need for a dedicated server / vps for most people now.

Image from the post
Image from the post

@maxjendrall Wait, this is super interesting. So it's remote, right? Can't be local, obviously.

Basically every ChatGPT Work customer's assistant gets its own laptop?

How and when did you hear about this?

7 views
In reply to @DeeJayMek · Sep 5, 2024

DJing for The Stone Roses last ever Irish show in Belfast. I played this gig with a broken wrist 🚑 Sound system ruled. Last 2 minutes of set⬇️
Kurtis Blow // Liquid Liquid // The Clash // Can // Talking Heads // Incredible Bongo Band // Ian Dury // Information Society // Hashim t.co/YQtYOP1zlD

Video from the post▶ Watch on Twitter

@DeeJayMek So much love for the Liquid Liquid x Clash moment. And, there's ALWAYS time for Jaki Liebezeit.

How did you get that gig?

29 views · 1 profile visit
In reply to @nicktozi · Feb 16, 2025

If you're waking up this morning grateful for another day...

You've already won, friend.

My life hack for this: a little golf counter, the clicker kind. I got a fancy one in silver and I carry it with me. When people say "count your blessings" I took that literally.

A few times a day I stop and ask what I could be thankful for, then click. Nice coffee, click, thank you. Easily 20-30 a day, from the mundane to the very high-spirited.

Today so far: my lovely mechanical keyboard, the guy who just delivered groceries, green tea.

What are you grateful for today?

4 views
In reply to @mehdizare · Sep 3, 2026

Claude shipped Fable 5.1. The feed will treat that as the model you should use now.

I looked at my own usage. The list-price meter was about $56k in 30 days. Memberships made it about $800. That is the gap the labs are selling: sticker vs seat.

I still use cheap or free models for rewrites and first drafts. OpenCode plus DeepSeek covers a lot of that. I save the expensive slow model for a long agent run, the same way I would not hire a specialist for a sore throat.

If the job is simple, use the simple tool.

Image from the post

@mehdizare Does this mean you actually have a $55k bill for your private work with Fable?

6 views
In reply to @0xAndros · Sep 9, 2025

crazy that X open-sourced their algo. i spent last night studying it and here are the main takeaways for viral posts:

1. Early, strong in-network engagement. Your followers are your first gatekeepers. Rapid likes, retweets, and replies in the first minutes signal quality. This helps you break out of your followers (also means you should definitely reply to comments)

2. High-quality engagement signals. The algorithm prioritizes high dwell time, sustained video watch time, and deep thread reads. Avoid anything that leads to quick bounces. Your content needs to be genuinely satisfying.

3. Good social proximity. Who engages with you matters. Tweets from authors close to you in RealGraph, or liked by your close contacts, get a significant boost. Strong local social proof is a powerful amplifier. Interactions from highly connected followers definitely help as well

4. Topic and cluster relevance at scale. For out-of-network reach, your content needs to align with large, active SimClusters (broad topics/communities). Niche tags won't get you broad distribution. This is where researching trending topics becomes critical. Original content also does MUCH better than retweets or replies

5. Freshness and sustained momentum. Recent tweets with ongoing engagement are heavily favored. The algorithm tracks "last engagement since creation" - it wants content that keeps generating interest over time, not just a one-off spike. Consistency is key.

Image from the post

Oh, this is really interesting. Glad I just found it again in my bookmarks.

Though: this is from September 2025, so it's about a year old now.

Have you kept monitoring it since? Would love to know which of these five still hold exactly as written, and which ones have quietly moved on you.

1 like · 22 views
In reply to @original_ngv · Sep 9, 2026

Please fail more in life. In college I lost 41 hackathons, failed 15 interviews and my startup was an epic flop.

During that time I pitched to VPs at Rakuten, scientists from NTRO, engineers from Microsoft and Google and even the government.

And today I work out of New York City with one of the finest CEOs in the world on AI projects that are changing people's lives every day.

I learnt more from my losses than my wins. My failures have given me more than my wins ever have.

So please everyone I beg of you. Fail more often. The view at the top comes from falling a lot of times and getting back up.

Love the specificity of the list. And yes, I learn more from losses than wins too. I've even proactively set myself up to fail just to learn faster.

The part I'd add: at some point it has to be more than just failing. You have to fail well, which really means you had to learn to learn first.

And that usually rests on some kind of post-failure safety net. Not necessarily rich parents. Mine was a curiosity safety net: lower middle class, but I have always loved learning early and kept getting food pellets from the universe for it. That's what made failing survivable for me.

So my ask to you is: looking back at those 41 hackathons and 15 interviews, what were your hidden enablers? Could be socioeconomic, could be intellectual, could be a person. Would love the nuance, because I think that's the part people can't see from the outside.

1 reply · 2 likes · 124 views

Was on the bike thinking about the whole mathematics/OpenAI thing.

Alan Watts: if we approached music the way we approach capitalism, the conductor who reaches the last note first would be the most successful. And the musician writing the shortest songs would win. Obviously silly. But is the journey is as important as the result?

Someone in a thread said a proof is only really valuable if you can explain it. I agree. Otherwise it's slop. Math slop?

But here's the conundrum: what if it works anyway? Throw brute force, money and machines at it, and out comes the actual record-shattering thing. Nobody can explain how.

Someone told me part of the payoff here would be physical prediction models weather forecasting for becoming orders of magnitude better. So you could argue: skip the drama, the back-and-forth, unlock the benefit for everyone, and let only the people whose job it is care about how we got there.

Basically, what if the slop solves world hunger?

And at the same time: the human version of this is years of grinding. You invest your time, your relationships, long stretches without reward. Even the million-dollar prize is nominal once you count opportunity cost.

These three, four, five researchers could have slid up the bell curve into a fund job and made that money anyway.

The most heartwarming part of the post that kicked all this off: the guy mostly said he shouldn't be credited, that it should go to his dude over there. Incredible sportsmanship at very high stakes.

Fast-forward 30 years. Are those researchers financially safe? Taken care of?

Risking your private time to hack on a problem for years IS risk. Spinning up 10.000 agents as part of a marketing budget is not at all risky.

Anyways, all of this is making me think harder about my own work and whether the techniques and methods I apply to my research are actually healthy ones.

322 views
In reply to @pmddomingos · Sep 9, 2026

Before the AI boom is over, we in science have a once-in-a-lifetime opportunity to get tech companies to blow millions on solving scientific problems they couldn’t care less about. Let’s not waste it. Hey OpenAI, rumor has it Anthropic is on the verge of cracking P = NP.

@pmddomingos "Couldn't care less about" is the part that sticks with me. Turns out solving a high-status problem doesn't absolve you of having to care about it. I believe that's new.

277 views · 1 profile visit
In reply to @paulg · Sep 10, 2026

This is why it's a good thing that American labs are in the lead. If Chinese labs were, their employees wouldn't be able to warn us about threats. But Chinese labs will encounter the same threats within a year, in complete silence. t.co/Bg5Vakbl4W

@paulg Hadn't thought about it from that angle. Maybe the underrated advantage of free speech isn't just the speech part but having a culture for civil conflict. You can point at something and say "this looks broken" and that can now be the start to find a solution or agreement.

1 reply · 1 like · 576 views · 2 profile visits
In reply to @MarioNawfal · Sep 9, 2026

Researchers spent 500 hours and $70,000 and found LG TVs logging plain text transcripts in standby, mapping every device in the house, and feeding it to LG's ad arm

Unplug the internet and it saves the files until you plug back in.

LG says its TVs don't record ambient conversations.

The evidence begs to differ.

216 million of these are sitting in living rooms.

Where are the regulators?

Writer: Daniel
t.co/5LLdbXOK1s

Oh! Saw this on HN a few days ago and it didn't land (probably got distracted by Navier-Stokes). The video did land though. Had to watch twice to get that the audio playing back was what the TV itself had recorded.

Would love to see the process (including the r00ting!), the on-device infra/setup and what gets sent where exactly.

Is there a repo or write-up?

1 repost · 2 likes · 336 views · 2 bookmarks · 3 profile visits
In reply to @ValerioCapraro · Sep 9, 2026

BREAKING: OpenAI might have stolen another major proof.

In a detailed Mastodon post, which I report in full in the comments, Andreas Thom presents several pieces of evidence suggesting that OpenAI may have trained Astra on conversations in which he and Gábor Kun were working on Gromov’s soficity conjecture, one of the ten problems OpenAI later announced Astra had solved.

I know Andreas. We met several times early in our careers. He is an exceptional mathematician, a leading expert on sofic and hyperlinear groups, and one of the most respected scholars in the field. He has spent two decades working on this problem.

If his account is correct, this is not a minor dispute over attribution. It would mean that unpublished human work was absorbed into a model and then presented to the world as a breakthrough by the model itself.

And if the allegations raised by Levent Alpöge, Tristan Buckmaster, and now Andreas Thom are all substantiated, we are no longer looking at isolated incidents.

We may be looking at one of the greatest intellectual scandals in the history of science.

AI is not discovering new mathematics.

AI is stealing human discovery.

Image from the post

Love that math is also moonlighting now as a tracer dye. Meaning: novel math ideas are made of such rare stuff (stuff being human curiosity, endurance, sweat, tears, sleepless nights, jamming with friends, etc.) that when a solution unexpectedly pops up in a more sterile, non-human environment, we're now learning to trace it back to its origins. My take here: Beyond the practical implications of the discovery itself, we're all gaining better tools to work critically with AI.

5 likes · 1.3K views · 1 bookmark · 6 profile visits
In reply to @Quasilocal · Sep 9, 2026

The way OpenAI handled this Navier--Stokes situation was the first time I've actually felt disgust at the whole AI in math situation. Like watching a billionare celebrate his hunting prowess by mounting the head of an endagered rhino on his wall, after 1000 soldiers trapped it.

@Quasilocal The debacle also had me looking at my own AI research critically, wondering if it contains a similar under-current of "gratuitous intellectual violence."

153 views
In reply to @forgebitz · Sep 9, 2026

i don't think tech people understand how little non-tech people care about tech

the average person does not have a basic understanding of how a computer works

the idea of everyone just prompting their own software is only for a small subset of people t.co/6fXWDcsRPU

@forgebitz Personally, I wouldn't use "tech" as a word to separate myself from "non-tech" people. Feels like a term that peaked in the early 2000s.

110 views · 1 bookmark · 1 profile visit
In reply to @BetterCallMedhi · Sep 9, 2026

anyone who thinks a tool can tell you if a text is AI written or not with 100% certainty is fucking retarded

these detectors are literally just checking text against basic metrics like perplexity/ sentence burstiness to spot predictable patterns BUT
modern LLMs are trained explicitly to match human language distributions…when your model's entire objective is to minimize the statistical gap between its outputs and actual human writing, the overlap becomes massive

that makes binary classification an absolute nightmare, sincce human and AI text distributions basically sit on top of each other now, pulling a threshold out of thin air guarantees a mountain of false positives &false negatives

Agree. I went down the rabbit hole from the other side: trying to prove a text was provably written by a human. Keyboard hardware events, breathing patterns via mic. Ambient light changes via camera and photosensor, even palm-rest temperature changes to detect hands.

All technically possible. All hackable with even a smidge of curiosity.

So now my litmus test is just: do I get the sense that the author cares? imho you can't fake that. I'm okay if they used AI for it.

4 views
In reply to @lisperati · Sep 9, 2026

I have to admit this was an impulsive shitpost to see if anyone would argue that it wouldn't be possible for me to make champagne, but some of you took it earnestly, sorry 😀

We don't really have grapes- though we do have cherries, loquats, apples, pears, and oranges t.co/RwGQnhaHn2

@lisperati Impulsive or not, I loved reading it. Fruit in a backyard. You, as a physical human, in an actual place (especially instead of all this late AI-headiness)

1 reply · 1 like · 28 views · 2 bookmarks · 1 profile visit
In reply to @emanperez28 · Sep 9, 2026

I’ve got to say I did not think a 16 second demo from my iphone would do crazy numbers. Just a reminder to use what you’ve got, no matter where you are.

@emanperez28 Love these demos. I tried something adjacent on macOS recently: detecting presence via the temperature rise where your palms rest on the keyboard. Got it working-ish but learned a lot about hidden sensors :)

2 likes · 69 views · 2 profile visits
In reply to @samagra_sharma · Sep 9, 2026

Introducing Whip Web: Youtube for interactive experiences

The next wave of internet culture will be something you step inside

Models like Astra are opening up a new creative frontier

Video found YouTube. Interactive experiences deserve a home too.

t.co/iQygo8a1qD

🧵 t.co/yEPVfjaPEO

Video from the post▶ Watch on Twitter

@samagra_sharma I'd love that future world. My favorite physical exhibition of the last decade was "L'Oublier et Accélérer" at the Galifet Foundation in Aix-en-Provence.

26 views · 2 profile visits
In reply to @signulll · Sep 9, 2026

the wearable category has converged on visible object attached to body which is kind of a failure if the primary job is passive sensing.

the ideal health wearable probably disappears entirely.

maybe we need to reinvent the ankle monitor.

@signulll I'd say the Oura Ring already solves this. Visible, yes, but subtle enough that it disappears. And beautiful should you choose to look at it.

2 likes · 54 views · 2 profile visits
In reply to @Ananyo · Sep 9, 2026

The animosity you see towards AI firms isn't technophobia. It stems from what looks like their asset-stripping of human knowledge and creativity in a pissing contest in pursuit of big IPO, heedless of undermining the people whose work their using... t.co/5NQfjv3JRA

@Ananyo Agreed, this isn't technophobia. It's the pissing contest toward some looming IPO, and people can feel that.Maybe the real value of this whole Navier-Stokes moment is that it's outlandish enough that we're finally noticing what's being strip-mined.

3 likes · 819 views · 1 bookmark · 7 profile visits
In reply to @BetterCallMedhi · Sep 9, 2026

behind the theatrical PR claim lies the crude reality of brute force test time compute scaling burning hundreds of billions of synthetic tokens across a massive inference flux to coordinate 10000ephemeral agents running recursive verification loops is not cognitive illumination it is the industrialization of exhaustive graph search where narrow reward functions force an army of digital checkers through a pre existing mathematical landscape

Very much agree with you here.

To me, already the initial (knee-jerk) reaction to the rumor to throw "tEn-ThOusAnd agents for EigHty-eIgHT hours" at it feels lame.

Maybe uninspired describes it better. And therefore uninspiring.

What I do know is: I definitely want to learn more about the guys who've been grinding on this on their own time for years now.

All of this even puts my OWN recent "Running the scientific method in an AI autoresearch loop" projects in a new, more critical light as well. Maybe they're lame, too!

Like Alan Watt's "The end is not the goal" quote:

[...] In music, though, one doesn’t make the end of the composition the point of the composition. If that were so, the best conductors would be those who played fastest.

Maybe I skipped to the end too fast, too?

Caveat 1: As a non-mathematician I don't understand the domain. So, what do I know?
Caveat 2: As a non-academic, I don't know if this "scooping" based on a rumor is common practice. Leaves a bad taste in my mouth, but again. What do I know?

1 reply · 1 like · 61 views · 2 bookmarks · 1 profile visit

Ever wonder "Why is my apartment so warm all the time?"

Turns out: my router has been quietly moonlighting 24/7 as a small space heater at 36℃/97℉.

And my fridge at 27℃/82℉. You can see how the hot air flows up, and makes sure my ceiling is nice and, err..., crispy.

I'm so glad I got this thermal phone adapter because I would NEVER even have considered these two as candidates. (I was 100% convinced the problem came from a broken floor-heating unit and/or wall-emission.)

Highly recommended!

PS: I now turn off the router at night. And yes, ambient temps have finally dropped to expected levels.

Image from the post
Image from the post