Show HN: Do Codex skills save tokens? A six-run task-size benchmark
Article URL: https://codex-howto-benchmark.nguyenvantamdk2.chatgpt.site Comments URL: https://news.ycombinator.com/item?id=49150997 Points: 2 # Comments: 0
Article URL: https://codex-howto-benchmark.nguyenvantamdk2.chatgpt.site Comments URL: https://news.ycombinator.com/item?id=49150997 Points: 2 # Comments: 0
I wanted the simplest possible AI meeting recorder that sees and hears your meetings, workflows etc. No cloud. No account. No subscription. Lumi a small CLI that uses the features already built into macOS. Everything stays on your Mac. Your recordings never leave your machine unless you choose to. It’s fully open source, so you can inspect exactly what it’s doing instead of trusting a black box. Simply run “lumi record start” once. If you’re looking for a low key, private alternative to cloud me
Article URL: https://aivisibility.pro/ Comments URL: https://news.ycombinator.com/item?id=49150752 Points: 2 # Comments: 0
Article URL: https://arxiv.org/abs/2607.25651 Comments URL: https://news.ycombinator.com/item?id=49150741 Points: 1 # Comments: 0
Ask HN: How accurate are AI face analyzers? Comments URL: https://news.ycombinator.com/item?id=49150702 Points: 2 # Comments: 0
<p> Cloud canvas where your team and AI agents works together </p> <p> <a href="https://www.producthunt.com/products/murmell?utm_campaign=producthunt-atom-posts-feed&utm_medium=rss-feed&utm_source=producthunt-atom-posts-feed">Discussion</a> | <a href="https://www.producthunt.com/r/p/1213367?app_id=339">Link</a> </p>
Put together an extensive open source test suite for voice assistants. You can view it including current leaderboard at: https://git.cicero.sh/aquila/ha-voice-test-suite/ Tests are reproduceable, with clear instructions on how to run them on your machine there. I only have a GPU with 4GB vRAM, hence only capable of testing the small LLMs like Qwen3 4B Instruct. Have Gemma 4 running right now, but it's insanely slow and probably another 24 - 48 hours before it finishes. Tried cloud models like Cl
<p> Local sandboxes for AI agents on your Mac, Linux, bare metal </p> <p> <a href="https://www.producthunt.com/products/npm-i-g-hotcell?utm_campaign=producthunt-atom-posts-feed&utm_medium=rss-feed&utm_source=producthunt-atom-posts-feed">Discussion</a> | <a href="https://www.producthunt.com/r/p/1213363?app_id=339">Link</a> </p>
Article URL: https://qwen.ai/blog?id=qwen3.8 Comments URL: https://news.ycombinator.com/item?id=49150470 Points: 307 # Comments: 113
Article URL: https://github.com/datascale-ai/opentalking Comments URL: https://news.ycombinator.com/item?id=49150333 Points: 1 # Comments: 0
"Critical evidence of unique, structured breeding habitats." ScienceAlert stories are written, fact-checked, and edited by humans, never generated by AI. Don't miss a story, subscribe here.

Article URL: https://github.com/gnt-ai/gnt Comments URL: https://news.ycombinator.com/item?id=49150247 Points: 3 # Comments: 2
Companies tend to pick their first agent job the same way. Somebody asks what AI could do for us, and the room converges on the work everyone can picture: write our blog posts, answer our customers, handle the inbox. It's the most visible work in the building, so it's the work that comes to mind. Six months later the pilot is quietly parked and the conclusion is that the technology wasn't ready. The technology was fine. The job selection was the problem, and visibility is what made it a bad…

<p><strong>Pershing Square Signature Center, New York</strong></p><p>John David Washington leads an exceptionally talented cast in a sleek off-Broadway thriller that feels a little dated</p><p>Making its US debut three years after its London premiere, Andrew Stein’s Disruption faces the typical battle of the topical play: even with the most expedient fast-track to production, works tackling current events are in danger of obsolescence upon arrival, often left at the mercy of the latest update, f

Preference alignment has become a crucial component in enhancing the performance of Large Language Models (LLMs), yet its impact in Multimodal Large Language Models (MLLMs) remains comparatively underexplored. Similar to language models, MLLMs for image understanding tasks encounter challenges like hallucination. In MLLMs, hallucination can occur not only by stating incorrect facts but also by producing responses that are inconsistent with the image content. A primary objective of alignment for
Circles uses the OpenAI API and Codex to power AI-native telco experiences, increasing ARPU by 22%, reducing churn by 9%, and improving development efficiency.
I've been using DeepSeek v4 flash for guided agent workflows (in-IDE, prompt/review diffs) and have found no noticeable benefit to using Haiku, Opus, Sonnet in this workflow. For around a dollar a day, I am able to more than double my own productivity. What are the frontier models for/what are they trying to solve? Are they designed to one shot applications? Is the expectation that we want to be able to "yolo" prompt LLMs and have them complete work unsupervised? Comments URL: https://news.ycomb
Article URL: https://atomadic.tech/orchard.html Comments URL: https://news.ycombinator.com/item?id=49149580 Points: 2 # Comments: 0
Article URL: https://github.com/spikonado/sprocket Comments URL: https://news.ycombinator.com/item?id=49149322 Points: 3 # Comments: 1
Article URL: https://blog.dlang.org/2026/06/07/teaching-an-ai-to-know-itself-building-a-local-llm-agent-in-d/ Comments URL: https://news.ycombinator.com/item?id=49149134 Points: 3 # Comments: 0
Article URL: https://github.com/micro/mu Comments URL: https://news.ycombinator.com/item?id=49148899 Points: 6 # Comments: 0
Get the 2026 AI Security & Cybersecurity Expert Bundle for $21 today only and access six courses covering AI security, ethical hacking, CompTIA prep, and penetration testing.

This code predicts the exact time and price of a future price point, and can also construct a curve of future prices. It can be used to decipher any process that has a graph—for example, a graph of mutual understanding between artificial intelligence and a human. Comments URL: https://news.ycombinator.com/item?id=49148735 Points: 2 # Comments: 0
The global memory chip shortage appears to be affecting the availability of Apple’s most popular Mac.
What is Decispher? The decision layer between your team and every AI agent Decispher is the system of record for engineering decisions. It automatically captures the decisions, conventions, constraints, and rationales your team produces every day from Slack, GitHub PRs, and docs, then serves that knowledge to both humans and AI agents the moment they need it. Think of it as a senior engineer who has read every Slack message, every PR, and every architecture discussion your team has ever had, and
Article URL: https://gavinray97.github.io/blog/design-by-contract-and-effects-for-llms Comments URL: https://news.ycombinator.com/item?id=49148299 Points: 2 # Comments: 0
On the latest episode of Equity, we discuss why Sam Altman has calling on the industry to "pace the rate of AI development."
Article URL: https://authoryze.ai Comments URL: https://news.ycombinator.com/item?id=49148057 Points: 3 # Comments: 2
<p> Make existing Codex and Claude Code sessions multiplayer </p> <p> <a href="https://www.producthunt.com/products/mpai?utm_campaign=producthunt-atom-posts-feed&utm_medium=rss-feed&utm_source=producthunt-atom-posts-feed">Discussion</a> | <a href="https://www.producthunt.com/r/p/1213234?app_id=339">Link</a> </p>
Article URL: https://github.com/paoloanzn/microcodex Comments URL: https://news.ycombinator.com/item?id=49147842 Points: 12 # Comments: 6
It changes over time. ScienceAlert stories are written, fact-checked, and edited by humans, never generated by AI. Don't miss a story, subscribe here.

Fender CEO Edward "Bud" Cole gave an interview to T3 in May celebrating the 75th anniversary of the Telecaster with comments on AI and music that initially flew under the radar. But it has started making the rounds recently, pouring more fuel on an already raging fire of bad PR following the company pissing off […]

A couple months ago, I was tired of keep copy-paste across different Agent Chatbot tabs, feeling like a slave for AI sessions, so I created Kota for my own work. Kota was designed around a concept of Knights of the Round Table. It treats human as part of a project team, and - Runs on top of existing CLIs with existing logins, - Gives each agent a persisted, viewable and editable identity, instead of sessions - Has a file-based long-term memory, - Uses a self-managed skillset, - Works on separate
Snapchat will stop recommending fully AI-generated Spotlight videos as more platforms crack down on low-quality AI content.

<p> Wild Pokémon appear while you wait for Claude Code </p> <p> <a href="https://www.producthunt.com/products/claudemon?utm_campaign=producthunt-atom-posts-feed&utm_medium=rss-feed&utm_source=producthunt-atom-posts-feed">Discussion</a> | <a href="https://www.producthunt.com/r/p/1213145?app_id=339">Link</a> </p>
Hello HN ..I am posting here because I have made this bad ass thing. I want to be clear. I'm trying to make somthing worthy of people buying it and getting good value, I have not really shown it off or had anyone look at my site past the occasional web crawler, and I did 100% make this. Its called 0verload and the things it does, I will now list. Automatcially tracks versions on all your source code. It dosnt replace GIT, its for the saves you do between checkins...that one version you should ha
<p> Study 25,000+ screens from top-earning iOS apps </p> <p> <a href="https://www.producthunt.com/products/appllama?utm_campaign=producthunt-atom-posts-feed&utm_medium=rss-feed&utm_source=producthunt-atom-posts-feed">Discussion</a> | <a href="https://www.producthunt.com/r/p/1213121?app_id=339">Link</a> </p>
I just need to know am i doing good or not for my future Comments URL: https://news.ycombinator.com/item?id=49145976 Points: 1 # Comments: 4
Welcome back to TechCrunch Mobility, your hub for the future of transportation and now, more than ever, the role AI is playing in it.