fix(anthropic): restore subscription OAuth billing routing
<pre style='white-space:pre-wrap;width:81ex'>fix(anthropic): restore subscription OAuth billing routing</pre>
<pre style='white-space:pre-wrap;width:81ex'>fix(anthropic): restore subscription OAuth billing routing</pre>
Article URL: https://github.com/makerchecker/MakerChecker Comments URL: https://news.ycombinator.com/item?id=48804182 Points: 11 # Comments: 5
Article URL: https://www.appblit.com/pdfreflow/demo Comments URL: https://news.ycombinator.com/item?id=48804162 Points: 4 # Comments: 0
I made a strategy game where you play the US or China through the AI race, 2026 to 2030, sixteen quarterly turns in the browser. One run takes about half an hour. At the start, the game seals two dice you never get to see. Inside: how hard alignment really is, and how fast takeoff compounds. You get eval reports, but only as ranges, and they flatter you most exactly when your systems are least aligned. At the end you get a debrief which shows what your evals said each quarter and also what was a
MicroPython (lexer, compiler and VM) on the SNES: 3.58 MHz 65816, 56 KB Python heap, 16-bit int. The REPL runs right inside the post via EmulatorJS and also works on real hardware via flashcart. I did this as a benchmark for Claude Fable. When the export ban hit, switching to Opus got the project stuck for three weeks; Fable came back and found the real bug in ninety minutes. Along the way: 23 compiler bugs and 4 MicroPython bugs, each root-caused with a minimal reproducer and filed upstream. If
A streaming box should not need a threat model. Neither should a username field, a demo repo, a reset flow, or a browser permission prompt. That is the irritating part this week: the risky pieces were ordinary. Home devices became a routing cover. Clean code pulled dirt from a dependency. Identity shortcuts aged badly. AI systems trusted the wrong instructions. Same soft spot throughout: trust

Station F, a Paris-based startup hub founded by French billionaire Xavier Niel, is gearing up for a new edition of its F/ai accelerator program in a bid to strengthen its positioning as a stepping stone for promising AI startups.
This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. South Korea’s hottest new bachelors are chip workers Baek, a 35-year-old manager at the South Korean semiconductor titan SK Hynix, was enrolled in a matchmaking company a year ago. In a…

Oracle's down more than 40% this month, the BIS thinks AI could destroy the economy, and we've got the Kettle on for a chat about the whole mess
Article URL: https://github.com/GrayboxTech/weightslab Comments URL: https://news.ycombinator.com/item?id=48803966 Points: 1 # Comments: 0
<pre style='white-space:pre-wrap;width:81ex'>fix(diagnostics-otel): route OTLP exports through env proxy * fix(diagnostics-otel): route OTLP exports through env proxy * fix(diagnostics-otel): harden OTLP proxy agent options * fix(ci): refresh rebased gateway checks * fix(ci): avoid proxy boundary scan</pre>
One thing I expected from the AI era was a decrease in emotional attachment to code reviews. My assumption was that AI-generated code would make reviews easier because criticism of the code wouldn't feel like criticism of the author. Instead, I've repeatedly seen people strongly defend AI generated code, and sometimes even defend AI explanations that are demonstrably incorrect. Is anyone else seeing this? If so, what do you think is driving it? How are you handling it within your team? Also, has
How I managed to automate building a KMP application Comments URL: https://news.ycombinator.com/item?id=48803807 Points: 1 # Comments: 0
Article URL: https://raheeljunaid.com/blog/anthropics-method-to-losing-goodwill-in-a-few-easy-steps/ Comments URL: https://news.ycombinator.com/item?id=48803751 Points: 235 # Comments: 181
Is this the beginning of a new world?

Nvidia's next AI server rack, Kyber NVL144, has been delayed more than a year to 2028 because of circuit board manufacturing problems, according to analyst firm SemiAnalysis. Asian suppliers lost up to double-digit percentages in market value. The more powerful Rubin Ultra variant has also been canceled. The setbacks could give AMD and Google an opening to compete. The article Nvidia's Kyber NVL144 reportedly pushed back more than a year, Asian suppliers drop appeared first on The Decoder .
Anthropic's Fable 5 promises mythic AI power, but surprise restrictions make me wonder if it's more trouble than it's worth for day-to-day use.
ByteDance and Alibaba are shutting down the features that let users build and chat with custom AI companions, responding to new regulations from Beijing. The article China forces its biggest AI platforms to shut down humanlike chatbot personas appeared first on The Decoder .

I hear more and more from people in companies of all sorts, are building internal tools/replacements of SaSS businesses. We built an automated QA agent based on Playwright, that does QA on our own products, literally in a week and we continue to contribute to it. Our QA engineers love it. What sort of things/products are you building? Comments URL: https://news.ycombinator.com/item?id=48803546 Points: 2 # Comments: 0
<p> Create your own Chrome Extensions by chatting with AI </p> <p> <a href="https://www.producthunt.com/products/plugthis-chrome-extension-generator?utm_campaign=producthunt-atom-posts-feed&utm_medium=rss-feed&utm_source=producthunt-atom-posts-feed">Discussion</a> | <a href="https://www.producthunt.com/r/p/1189320?app_id=339">Link</a> </p>
This article looks at five concrete ways SLMs are showing up inside next-generation agents right now, from the research backing them to the tools and numbers worth knowing if you're deciding whether your next agent needs a frontier model at all

<p>These systems will soon be able to track our public and private lives. But we can make the policy choices to reject it</p><p>In the near future, <a href="https://www.theguardian.com/technology/artificialintelligenceai">AI</a>-powered surveillance systems will be able to track everything we do in public, and much of what we do in private. And if we do something wrong – shoplift, litter, jaywalk, you name it – the system will notice, retain it, tie it to your official government record, communi

"I never wake up in a bad mood." ScienceAlert stories are written, fact-checked, and edited by humans, never generated by AI. Don't miss a story, subscribe here.

Best-worst comparisons, MaxDiff-style judging, and Plackett-Luce utility scores give agent teams a cleaner way to decide which configs to ship, prune, and route toward next. The post Stop Ranking Agent Configs by Average Score appeared first on Towards Data Science .
<!-- SC_OFF --><div class="md"><p>Today I was just casually browsing some jobs with tags [machine learning] on one of those large popular job-sites. What I am seeing really had me astonished. I want to check with Reddit whether I am hallucinating.</p> <p>A non-FAANG/non-Deepmind/.../non-Anthropic industrial automation company is hiring people to work on ML for robotics. Fine. But then I saw their laundry list of job requirements ("you must meet these"), which include:</p> <ul> <li>Deep expertise
Comments URL: https://news.ycombinator.com/item?id=48803415 Points: 5 # Comments: 0
I'm looking for this myself but figured it's good to have an actual discussion about this. I'm pretty new to the benchmarking side of LLMs. For example, I looked at eyeballvull [1]. It seems promising but I don't see wide support for example. I respect the author for still committing. A benchmark where an agent scans a repo in full is what I'm looking for. But then I also wondered: maybe there are others out there that I haven't been aware of yet. And, at the risk of potentially being flooded, f
You build an agent with five tools.

Building a shortlist for an AI SOC evaluation can be tough. SIEM, SOAR, and pureplay AI SOC vendors are all saying the same thing. But behind the identical label sit very different products, from chat assistants bolted onto a legacy SIEM to agent platforms that run detection, triage, investigation, and response on their own data foundation. Whether a platform will materially change outcomes for


Amazon Web Services is shutting down its crowdsourcing service Mechanical Turk to new customers starting July 30, 2026. The article Amazon sunsets Mechanical Turk, the original "Artificial Artificial Intelligence" appeared first on The Decoder .

I kept hitting context window loss on long agent runs and realized the problem was prompt bloat, not model size. Comments URL: https://news.ycombinator.com/item?id=48803124 Points: 2 # Comments: 0
<p>FCA’s review into how tech will reshape financial services warns about amplified risks of cyber-crime and fraud</p><ul><li><p><a href="https://www.theguardian.com/business/live/2026/jul/06/sky-takeover-itv-broadcasting-media-deal-business-live-news">Business live – latest updates</a></p></li></ul><p>Ministers have been urged to toughen the City regulator’s powers to protect consumers against the potential risks of AI, according to a landmark review.</p><p>The Mills review by the Financial Con

<p>Researchers say small changes in drafting could spread rapidly and create long-term shifts in public opinion</p><p>AI tools are twisting online messages on sensitive political topics about everything from abortion to climate change in ways that could snowball to reshape long-term public opinion, experts have said.</p><p>As tech companies push AI tools as convenient ways to redraft and summarise the massive influx of daily messages, many inject their own political biases – some leaning distinc

An AI companion sounds dystopian, but it has become a common thread in the wider conversation about the perils of generative AI. What it refers to is essentially a conversational agent built to sustain an ongoing, personal relationship with a user, with the memory and steady persona that keep it consistent from one session to […] The post China’s AI companion rules: what Beijing is really going after appeared first on AI News .

France is accelerating its transition to post-quantum encryption: France’s cybersecurity agency ANSSI said on Tuesday it would stop certifying security products that lack quantum-resistant encryption, a move that will force government bodies and critical operators to shift away from older systems. Samih Souissi, ANSSI’s chief of staff, said at the France Quantum conference that the agency would halt such certifications from 2027, and that businesses should be buying only quantum-safe products by
Article URL: https://bestjournalapp.com Comments URL: https://news.ycombinator.com/item?id=48802890 Points: 2 # Comments: 1
Hi HN! I've been spending the past month on PAI, a Linux-esque harness for a personal assistant AI for your Mac. Instead of giving a chatbot a bunch of tools, I wanted to just drop the LLM into a terminal and see how it'd fare. I've been using it to manage my calendar, 5+ email accounts, and engaging with / responding to people in a timely manner (casus belli for this entire project). After seeing how other always-on assistants were designed, I thought the Unix / Linux philosophy of system desig