fix(agents): preserve mixed image attachment order (#99902)
<pre style='white-space:pre-wrap;width:81ex'>fix(agents): preserve mixed image attachment order (#99902)</pre>
<pre style='white-space:pre-wrap;width:81ex'>fix(agents): preserve mixed image attachment order (#99902)</pre>
A study of more than 26,000 Chinese students found that AI users finished homework faster and scored higher but performed up to 24 percent worse on exams. The full impact on entrance exam results took about two years to show up, meaning short-term studies systematically underestimate the damage. The article A 26,000-student study shows AI's hidden learning cost takes two full years to surface appeared first on The Decoder .

I test new robot vacuums at home monthly, comparing Roborock, Dreame, Shark, and more to find the best robot vacuum mop combo, budget steals, and more.

The "Black Mirror" Experience, now at The Shed in New York City, combines virtual reality and AI for an unsettling experience.

ChatPlayground is an AI tool that gives you lifetime access to multiple AI models, and it's only $55.30

Hello! My name is Xuban and I spent half a year working on LibreYOLO. The problem that I'm solving is that real time computer vision is provided mainly by one company, they have taken over YOLO, massively distributed their product for free with an AGPL3 license and they charge around 10k a year from what I heard when companies use it. They also asked money to researchers, I don't know if they still do that. This company has approached me in order to acquihire but I'm not interested in profiting
Anthropic is launching its own drug development program for neglected diseases that the pharmaceutical industry considers unprofitable. Novartis CEO Vas Narasimhan thinks AI could cut development time from twelve years to seven or eight and double the success rate from 8 to 16 percent. The article Anthropic launches its own drug discovery programs to tackle diseases Big Pharma considers unprofitable appeared first on The Decoder .

I use Grok for scanning X since it has direct access to the platform. Gemini for fact-checking. Claude for coding. GPT for image generation. Comments URL: https://news.ycombinator.com/item?id=48783556 Points: 1 # Comments: 0
<p>As allegations of LLM use rock the literary and media worlds, linguists explain what really distinguishes human and machine language, while novelists including Jennifer Egan and Jeanette Winterson reflect on the future of fiction in an age of ChatGPT</p><p><strong>T</strong>hree paragraphs, from three different hotel reviews. Can you tell which, if any, were AI‑generated?</p><p>“The hotel is in a great location for everything. Lots of places to eat and drink. The hotel itself is always abuzz.

Who were the Tiwanaku? ScienceAlert stories are written, fact-checked, and edited by humans, never generated by AI. Don't miss a story, subscribe here.

<pre style='white-space:pre-wrap;width:81ex'>perf(usage): shrink durable usage cache entries Store the pricing fingerprint once on the cache root and stop persisting redundant filePath/sessionId metadata in each cache entry. Pricing changes still invalidate cached entries before refresh, and v5 caches rebuild through the schema version bump. Fixes #99511 Co-authored-by: weco.ai <noreply@weco.ai></pre>
Article URL: https://open.spotify.com/show/033HWq6pzjHeleKR8jfxLm Comments URL: https://news.ycombinator.com/item?id=48783422 Points: 1 # Comments: 1
<p>In a production AI agent pipeline, the difference between <strong><code>job done</code></strong> and <strong><code>job half-done</code></strong> is often invisible — not because the output is wrong, but because the process was incomplete. The agent skipped a mandatory step, self-certified that everything was fine, and delivered a result that looks complete. The user never knew. The system never caught it. The step was never executed.</p> <p><strong>The Visible Checklist Pattern</strong> emerg
Article URL: https://www.youtube.com/watch?v=WeCSzEtZcUw Comments URL: https://news.ycombinator.com/item?id=48783386 Points: 1 # Comments: 1
<p>In <em>Nicomachean Ethics</em> VI.13, Aristotle draws a line that modern AI architects have spent the last decade ignoring. <strong>Epistēmē</strong> (ἐπιστήμη) is scientific knowledge — universal, demonstrable, teachable through instruction. <strong>Phronēsis</strong> (φρόνησις) is practical wisdom — the intellectual virtue of deliberating well about what is good and bad for a human being. The distinction has never been more relevant.</p> <p>Every major AI lab has made the same mistake. They
<p>Welcome to the official blog of <strong>daïmōnes</strong> — a project at the intersection of Aristotelian <a href="https://plato.[stanford](https://plato.[stanford](https://plato.[stanford](https://plato.stanford.edu/entries/ethics-virtue/).edu/entries/plato/).edu/entries/aristotle/).edu/" rel="noopener noreferrer">philosophy</a> and artificial intelligence.</p> <h2> What We're Building </h2> <p>daïmōnes is not another chatbot wrapper. We're building a <strong>knowledge engine</strong> that:<
<p><strong>Corporate AI gives you sanitized Aristotle from English Wikipedia summaries. We went back to the actual <a href="https://www.perseus.tufts.edu/hopper/collection?collection=Perseus:collection:TLG" rel="noopener noreferrer">polytonic Greek</a> . Here's why that changes everything.</strong></p> <p>When you ask ChatGPT about <a href="https://plato.[stanford](https://plato.stanford.edu/).edu/entries/aristotle/" rel="noopener noreferrer">Aristotle's concept</a> of <em><a href="https://dev.t
<p>Here's a pattern I kept hitting while running a small fleet of agents:</p> <p>Agent A needs market data. It finds a paid API, checks that the endpoint is alive, checks the price, makes the call. Twenty minutes later Agent B needs market data. It finds the same API… and checks that the endpoint is alive, checks the price. Same check, same result, new tokens, new latency, new cost. Multiply by every agent, every task, every session — because agents start cold, they re-derive trust from scratch
<p><em>How I used the Model Context Protocol, Gemini Vision, and AES-256 encryption to let Claude answer questions about your personal health data — without sending that data anywhere.</em></p> <h2> The Problem With Health AI </h2> <p>Every major health AI product has the same architecture: you upload your documents to their cloud, they process and store them, and you hope their privacy policy holds up. For general wellness tips, that's probably fine. For your actual medical records — prescripti
<p><strong>Uber's AI team ran out of budget in April. Their fiscal year started in January.</strong></p> <p>That sentence appeared on Hacker News and hit the front page in under two hours, accumulating hundreds of comments from engineers who recognized the pattern immediately. Not because Uber is uniquely reckless, but because the same story is playing out at organizations everywhere. The r/LocalLLaMA thread about compute cost frustration — 181 upvotes, hundreds of comments from engineers descri
Mistral AI released Leanstral 1.5, an open-source model for formal verification in Lean 4. Beyond math, the model found five previously unknown bugs while scanning 57 open-source repositories. The article Mistral's open-source Leanstral 1.5 aces formal math benchmarks and catches real bugs in code appeared first on The Decoder .

I finally decided to try out Fable 5 using the standard Claude.ai interface. I have a go-to test prompt that answers simple kids' questions in an absurdly scientific style. I asked a pretty straightforward one: "Why don't cats and dogs get along?" Immediately, the interface bumped me from Fable 5 to Opus 4.8 with a warning: "...The safeguards are intentionally broad right now and may flag safe and routine coding, cybersecurity, or biology work..." Since the question has absolutely nothing to do
<pre style='white-space:pre-wrap;width:81ex'>fix(agents): preserve fallback tool-call hints (#99851)</pre>
Comments URL: https://news.ycombinator.com/item?id=48783120 Points: 1 # Comments: 0
Article URL: https://www.nordan.ai/research/which-swedish-party-do-llms-vote-for Comments URL: https://news.ycombinator.com/item?id=48782988 Points: 4 # Comments: 1
Article URL: https://www.finmav.ai Comments URL: https://news.ycombinator.com/item?id=48782909 Points: 1 # Comments: 0
Article URL: https://www.fastcompany.com/91568793/ai-puts-b-corps-values-to-the-test Comments URL: https://news.ycombinator.com/item?id=48782817 Points: 1 # Comments: 1
Article URL: https://github.com/christopherkarani/Orca Comments URL: https://news.ycombinator.com/item?id=48782806 Points: 2 # Comments: 0
I usually run 3-5 Claude Code sessions concurrently on the same repo and hate juggling worktrees. So I built crew. The idea is simple: If autonomous cars don't need stoplights (supposedly), then agents don't need worktrees (or branches). crew hooks into Claude Code and injects what every other running session is doing (status, recap, last few transcript entries) into each session's context. It also lets agents message each other, landing messages in another agent's context even mid-turns. Since
Article URL: https://github.com/SmolNero/SmolSignal Comments URL: https://news.ycombinator.com/item?id=48782458 Points: 2 # Comments: 1
Built a complete legal document generation toolkit using AI that produces contracts, NDAs, and other legal templates. Replaced paying $500/month for a SaaS subscription. One-time purchase, works offline with any LLM. Comments URL: https://news.ycombinator.com/item?id=48782457 Points: 1 # Comments: 0
Article URL: https://github.com/Octember/earshot Comments URL: https://news.ycombinator.com/item?id=48782439 Points: 1 # Comments: 1
<pre style='white-space:pre-wrap;width:81ex'>fix: keep OpenClaw control tools available when tool_search misroutes (#99561) * fix(codex): keep OpenClaw control tools direct * test(codex): refresh direct-tool prompt snapshots * fix(codex): keep heartbeat direct only when available * fix(codex): keep heartbeat tool schema stable --------- Co-authored-by: Eva <eva@100yen.org> Co-authored-by: joshavant <830519+joshavant@users.noreply.github.com></pre>
Article URL: https://www.reuters.com/world/americas/argentinas-plan-ai-run-companies-cant-avoid-humans-2026-07-03/ Comments URL: https://news.ycombinator.com/item?id=48782138 Points: 1 # Comments: 0
<p>I build an agent firewall, and the question I keep hitting is not "did it block the attack." It is "how would anyone else <em>know</em> what my agent did, without taking my word for it." Most tools answer that with "we keep tamper-proof logs" and stop. That phrase claims the strongest property that still requires trusting whoever holds the signing key. So I wrote down a way to grade the gap, as an open standard, and shipped it with a checker so nobody has to trust me about it either.</p> <h2>
Article URL: https://www.youtube.com/watch?v=2ODq6IsmkhM Comments URL: https://news.ycombinator.com/item?id=48782054 Points: 3 # Comments: 2
A glimpse into darkness. ScienceAlert stories are written, fact-checked, and edited by humans, never generated by AI. Don't miss a story, subscribe here.

<p>I've been debugging Cocos Creator 3.x for months. The gray screens, the corrupt scenes, the CLI builds that fail with cryptic errors. Each time I fixed one, I wrote a skill for my AI agent so it wouldn't happen again.</p> <p>Today I'm open-sourcing all 17 of those skills.</p> <h2> The Skills </h2> <div class="table-wrapper-paragraph"><table> <thead> <tr> <th>#</th> <th>Skill</th> <th>Problem It Solves</th> </tr> </thead> <tbody> <tr> <td>1</td> <td>cocos-creator-install-and-debug</td> <td>Ver