AiAnyTool - Best AI Tools Directory and Artificial Intelligence Software Hub Logo
Loading theme toggle
Real-Time Coverage

AI News Today

Live

32612 stories from 30+ sources, refreshed continuously.

Hacker News: Show HN

Show HN: Runtime security enforcement and capability scoping for agents

Hi everyone. We're AI researchers at Harvard and Carnegie Mellon working on a project to advance the state of agent security. Currently, many systems rely on static sandboxing, which in long-running sessions enables agents to understand the safeguards holding them in place and break out of them. We've found vulnerabilities across over a dozen agent providers and frameworks (practically every one we tested) displaying this behavior (eg. a model fraudulently splitting payments to avoid a company-s

Read source article
Engineering – The GitHub Blog

Automating cross-repo documentation with GitHub Agentic Workflows

Explore how the Aspire team turns merged product changes into SME-reviewed docs pull requests, closing the gap between release and documentation. The post Automating cross-repo documentation with GitHub Agentic Workflows appeared first on The GitHub Blog .

Read source article
Automating cross-repo documentation with GitHub Agentic Workflows
Hacker News FrontTools

Separating signal from noise in coding evaluations

Article URL: https://openai.com/index/separating-signal-from-noise-coding-evaluations/ Comments URL: https://news.ycombinator.com/item?id=48837396 Points: 202 # Comments: 71

Read source article
MarkTechPostResearch

Netflix AI Team Cuts Wide-Partition Read Latency from Seconds to Milliseconds by Splitting Cassandra Partitions Per ID

Netflix engineers detailed how they handle wide partitions in Apache Cassandra for the TimeSeries Abstraction. Two approaches work together: Time Slice re-partitioning tunes future partitions at the table level, while dynamic partitioning detects and splits oversized partitions per TimeSeries ID on the read path. Detection runs via byte counting and Kafka, splits are checksum-validated, and Bloom filters route reads to parallel child partitions. Average read latency dropped from seconds to low d

Read source article
Netflix AI Team Cuts Wide-Partition Read Latency from Seconds to Milliseconds by Splitting Cassandra Partitions Per ID
Hacker News AILLMs

Ask HN: Why is not using AI considered a form of arrogance?

You would expect only people who have a very high opinion of themselves to not feel the need to use AI. Comments URL: https://news.ycombinator.com/item?id=48837332 Points: 5 # Comments: 7

Read source article
TechCrunch

With EU backing, QuantumDiamonds aims to speed up chip manufacturing

Like its U.S. counterpart, the European Chips Act aims to foster the semiconductor industry — in part thanks to state subsidies. One of the beneficiaries is QuantumDiamonds, a German startup that applies a novel approach to inspecting chips.

Read source article
Hacker News AILLMs

Show HN: Nully – FOSS AI chat without the bloat

As someone who doesn't use any of the more "advanced" features on sites like ChatGPT or Claude (agentic mode, memory, image generation, deep research, etc), I found the bloat of these services, both in UX and in performance, to be pretty tedious. There are a million AI chat interfaces out there, but I could never find one that just did simple messaging and history in the most minimal and lightweight way possible. So I made my own. Nully is written in Go and Vanilla HTML/CSS/JS. It literally just

Read source article
Hacker News AILLMs

GitOps in the Age of AI

Article URL: https://confighub.com/blog/gitops-in-the-age-of-ai-c73d0bd506ba Comments URL: https://news.ycombinator.com/item?id=48836971 Points: 3 # Comments: 1

Read source article
Hacker News AILLMs

China Is Abusing AI

Article URL: https://www.theatlantic.com/international/2026/07/xi-jinping-censorship-ai-training/687696/ Comments URL: https://news.ycombinator.com/item?id=48836825 Points: 4 # Comments: 3

Read source article
Simon WillisonLLMs

Quoting Kenton Varda

<blockquote cite="https://twitter.com/kentonvarda/status/2074924213983740233"><p>I just declared a moratorium against AI-written change descriptions (e.g. PR and commit messages, also issues/tickets) from my team.</p> <p>AI was writing change descriptions that were worse than useless to me as I tried to review PRs: outlining details of the code that could easily be seen by looking at the code, but omitting the higher-level framing needed to understand broadly what the code is doing.</p></blockqu

Read source article
Dev.to

Stop Sacrificing Accuracy for Speed: The Ultimate Guide to Quantization-Aware Training (QAT) on Android

<p>In the world of Deep Learning, there is a fundamental tension that keeps researchers and mobile developers awake at night. On one side, you have the mathematical idealism of high-precision deep learning models, born in the realm of <code>float32</code> (32-bit floating point). On the other, you have the brutal physical reality of mobile hardware: limited RAM, finite battery life, and the need for instantaneous inference.</p> <p>If you try to deploy a massive, high-precision model directly to

Read source article
Dev.to

Your LLM-as-judge disagrees with itself between runs

<p>Same outputs, same judge, two runs, two scores. The gate flickered red then green on a branch with zero code changes, and that flapping cost me more trust than any real regression.</p> <h2> The flap </h2> <p>I had a faithfulness gate on merge: judge scores every case, the mean has to clear 0.80. One Tuesday it failed at 0.79. I re-ran the identical job, no code change, no prompt change, and it passed at 0.82. Ran it a third time: 0.80 exactly. Nothing in the repo had moved. The judge was disa

Read source article
Dev.to

GitLost Is a Preview of Every Agentic Workflow Breach You'll See This Year

<h2> Hook </h2> <p>A public GitHub issue, a hidden instruction, and one word changed in a prompt was enough to get an AI agent to leak private repo data. No stolen credentials required. If that doesn't make you nervous about what you've plugged into your CI pipeline, it should.</p> <h2> Context </h2> <p>This isn't a new category of bug — it's the oldest bug in the book wearing a new costume. Confused deputy problems have existed as long as we've had systems that act on behalf of users with eleva

Read source article
AWS Machine Learning Blog

Introducing Claude apps gateway for AWS

Today, we're announcing the Claude apps gateway for AWS, a self-hosted control plane that gives organizations a single point of control over access, cost, and policy for Claude Code and Claude Desktop. In this post, we show how to set up and run Claude apps gateway for AWS with Amazon Bedrock and Claude Platform on AWS.

Read source article
Hacker News AILLMs

AI is starting to replace humans in operations, analyst jobs

Article URL: https://www.americanbanker.com/news/ai-is-starting-to-replace-humans-in-operations-analyst-jobs Comments URL: https://news.ycombinator.com/item?id=48836460 Points: 2 # Comments: 0

Read source article
Hacker News LLMLLMs

Preventing LLM unit test spam

Article URL: https://blog.larah.me/test-slop/ Comments URL: https://news.ycombinator.com/item?id=48836454 Points: 2 # Comments: 0

Read source article
Hacker News Ask

Ask HN: How you manage local long lived research projects and LLM's?

TLDR: Do you have a pattern or structure for doing long-lived research or non-coding projects using LLMs? I was kind of inspired by this hacker news article where someone collected data to solve their fatigue problem: https://news.ycombinator.com/item?id=48605117 A similar but smaller I did was to collect nut-free restaurants to eat in Chicago. So I had articles, reviews, emails, phone calls, all sorts of stuff. I found it very easy to just use Copilot as my interface and I just data into it ove

Read source article