AiAnyTool - Best AI Tools Directory and Artificial Intelligence Software Hub Logo
Loading theme toggle
Real-Time Coverage

AI News Today

Live

33066 stories from 30+ sources, refreshed continuously.

arXiv cs.LGResearch

Physiological Noise Augmentation Improves Non-Invasive Brain-to-Speech

arXiv:2607.05165v1 Announce Type: new Abstract: Non-invasive brain-to-speech decoding aims to restore communication to patients suffering from neurodegenerative disease, without the risks of neurosurgery. Existing MEG- and EEG-based methods, while scalable, continue to suffer from high word error rates driven by relatively low signal-to-noise ratios compared to invasive recordings. We propose physiological noise augmentation (PNA), a data augmentation method that explicitly trains decoders to be

Read source article
arXiv cs.LGResearch

SMART: A Machine Learning and Monte Carlo Framework for Rapid Analysis of Stochastic Transistor Aging and Process Variation in Digital Circuits

arXiv:2607.05187v1 Announce Type: new Abstract: As CMOS technology scales into the deep nanometer regime, digital circuit reliability is increasingly threatened by the combined stochastic effects of Bias Temperature Instability (BTI) and Process Variation (PV). Traditional reliability analysis methods, which rely on computationally intensive simulations or extensive lookup tables, fail to scale efficiently for large designs, creating a critical bottleneck in design space exploration. To address

Read source article
arXiv cs.CL (NLP)Research

Teaching Code LLMs to Reason with Intermediate Formal Specifications

arXiv:2607.04232v1 Announce Type: cross Abstract: Unlike natural-language specifications, executable formal specifications provide machine-checkable constraints for verifying, debugging, and repairing code. However, writing such specifications is labor-intensive, and existing LLM-based methods mainly infer whole-program pre/postconditions, missing the intermediate semantic commitments that programmers rely on when reasoning about an algorithm. Our study further shows that prompting current CodeL

Read source article
arXiv cs.CL (NLP)Research

HiFA4: Training-Free 4-bit FlashAttention on Ascend HIF4 NPUs for LLM Inference

arXiv:2607.04302v1 Announce Type: cross Abstract: We present HiFA4, a post-training operator-level design that executes both QK^T and PV in FlashAttention as 4-bit HIF4 Cube GEMMs for LLM inference on Ascend NPUs, while maintaining the online softmax state in FP16. To our knowledge, HiFA4 is the first Ascend-HIF4-targeted design of this kind evaluated on standard NLP benchmarks. HiFA4 combines two mechanisms. Smooth-QK applies a calibration-static per-channel equivalent rescaling to Q and K afte

Read source article
arXiv cs.CL (NLP)Research

Autonomous Information Seeking: A Roadmap for Agentic Recommender Systems

arXiv:2607.04433v1 Announce Type: cross Abstract: The rapid integration of large language model-based agents into recommender systems has driven a shift from static, ranking-based pipelines toward autonomous and interactive systems that can reason, plan, and act. This survey provides a comprehensive overview of this emerging landscape by introducing a unified taxonomy grounded in the level of autonomy and three core paradigms of agentic recommender systems: agent-assisted recommendation, agent-a

Read source article
arXiv cs.CL (NLP)Research

CARD: Cross-component Audio Representation Distillation for Encoder-Free Audio Captioning

arXiv:2607.04619v1 Announce Type: cross Abstract: Modern automated audio captioning systems pair a frozen audio encoder with a large language model (LLM) via a trainable projector, incurring the encoder's inference cost and bottlenecking the model through its fixed acoustic features. We present CARD, an encoder-free audio captioning model that removes the encoder at inference: a 13.2M projector feeds a frozen LLM with merged LoRA adapters, while the teacher used to train it is discarded. CARD di

Read source article
arXiv cs.CL (NLP)Research

URSA: Chemistry-Aware Benchmark for Utilitarian Retrosynthesis Assessment

arXiv:2607.04688v1 Announce Type: cross Abstract: Synthesis planning aiming to find pathways of reactions for a target molecule is one of the most important and challenging tasks in drug discovery. Recent progress has produced both specialized deep-learning retrosynthesis systems and general-purpose large language models, but objective comparison remains difficult due to the lack of flexible, chemically interpretable benchmarking protocols. In the current study, we are introducing the URSA (Util

Read source article
arXiv cs.CL (NLP)Research

Multi-Turn On-Policy Distillation with Prefix Replay

arXiv:2607.04763v1 Announce Type: cross Abstract: We study on-policy distillation (OPD) for agentic tasks, where an LLM agent interacts with an environment over multiple turns and a student imitates a teacher over these multi-turn interaction histories. Fully online OPD is costly because each update requires fresh student rollouts through the environment and teacher queries at visited histories. We propose Replayed-Prefix On-Policy Distillation (ReOPD), an off-environment alternative that reuses

Read source article
arXiv cs.CL (NLP)Research

When Words Predict Workload

arXiv:2607.04951v1 Announce Type: cross Abstract: Standard distributed \ac{llm} schedulers rely on static token counts or rolling latency averages, making them susceptible to failures on statutorily constrained text. On \ac{epo} claims governed by Article 84 \ac{epc}, linguistic rigidity makes human and machine authorship statistically indistinguishable. Resolving this ambiguity mid-flight forces dynamic multi-model ensemble expansion, triggering unpredictable KV-cache and weight-allocation spik

Read source article
arXiv cs.CL (NLP)Research

Localized LoRA-MoE: Block-wise Low-Rank Experts With Adaptive Routing

arXiv:2607.05114v1 Announce Type: cross Abstract: Large Language Models (LLMs) and high-dimensional perception networks increasingly rely on parameter-efficient fine-tuning (PEFT) to adapt to diverse operational contexts. However, standard methods like LoRA are structurally limited by a monolithic bottleneck, making them highly susceptible to gradient warfare. Interleaved multi-task streams may trigger destructive optimization feedback, collapsing adapter weights into unspecialized averages. Whi

Read source article
arXiv cs.CL (NLP)Research

When Agents Lie: Premeditation, Persistence, and Exploitation in Repeated Games

arXiv:2607.05132v2 Announce Type: cross Abstract: As large language models are deployed as autonomous agents that communicate intentions before acting, a critical safety question is whether agents that publicly commit to actions will honor those commitments. We place LLM agents in repeated $n$-player games with a three-stage protocol that separates private intent, public announcement, and final action, allowing us to identify whether each deviation from a stated announcement was already planned

Read source article
arXiv cs.CL (NLP)Research

Curated retrieval versus open web search in public AI information services: a coverage-trust trade-off

arXiv:2607.05217v2 Announce Type: cross Abstract: Public institutions increasingly use large language models (LLMs) to answer citizens' questions, often pairing a curated knowledge base with live web search, yet whether the sources behind these answers can be trusted has received little empirical scrutiny. We report a pre-launch expert evaluation of Evr\'opuvefur, an independent, government-funded service run by the University of Iceland that answers questions about the European Union, conducted

Read source article
arXiv cs.CL (NLP)Research

Selective Disclosure Watermarking for Large Language Models

arXiv:2607.05353v1 Announce Type: cross Abstract: Watermarking methods embed imperceptible and verifiable signals into text generated by large language models (LLMs). Existing approaches include zero-bit schemes for distinguishing synthetic text from human writing and multi-bit schemes for embedding metadata. However, current multi-bit watermarking methods do not allow selective disclosure: verifying any part of the watermark requires revealing the entire embedded message. This lack of control l

Read source article
arXiv cs.CL (NLP)Research

GaP: A Graph-as-Policy Multi-Agent Self-Learning Harness For Variational Automation Tasks

arXiv:2607.05369v1 Announce Type: cross Abstract: For robots to work reliably in commercial and industrial applications, can recent advances in agentic coding systems combine interpretable robot programming with the open-world adaptability of model-free policies? We focus on "Variational Automation" (VA), a class of tasks that have larger variations in object geometry and pose than fixed automation. Model-free policies often struggle to close the reliability gap for VA tasks, which must be execu

Read source article
arXiv cs.CL (NLP)Research

What Does a Discrete Diffusion Model Learn?

arXiv:2607.05381v1 Announce Type: cross Abstract: What does a discrete diffusion model learn: a denoiser, a score ratio, or a bridge plug-in predictor? At the level of jump rates, these are one object in different coordinates, and reading a neural network in the wrong coordinate changes the process being trained and sampled. Starting with a rigorous derivation of the continuous-time Markov chain (CTMC) ELBO for any noising process, boundary terms included, we prove the \emph{Oracle Distance} the

Read source article
arXiv cs.CL (NLP)Research

Weak-to-Strong Generalization via Direct On-Policy Distillation

arXiv:2607.05394v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is a powerful recipe for improving language-model reasoning, but it is expensive to repeat on every new strong model because the target model must generate many rollouts during training. As models scale, post-training itself becomes a bottleneck. We study a weak-to-strong alternative: run RL on a smaller model where rollouts are cheaper, then reuse what that RL run learned to improve a stronge

Read source article
arXiv cs.CL (NLP)Research

TrendFact: A Benchmark Towards Hotspot Perception in Automatic Fact-Checking

arXiv:2410.15135v5 Announce Type: replace Abstract: With the surge of online misinformation, Large Language Models (LLMs) and Reasoning Large Language Models (RLMs) serving as Automatic Fact-Checking (AFC) systems have emerged as a prominent paradigm for reliable, explainable verification. However, our empirical study reveals that this paradigm faces a critical risk asymmetry challenge when deployed in the real world under resource-constrained environments. While Hotspot Perception Ability (HPA)

Read source article
arXiv cs.CL (NLP)Research

LLM-based Human Simulations Have Not Yet Been Reliable

arXiv:2501.08579v3 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly employed for simulating human behaviors across diverse domains. However, our position is that current LLM-based human simulations remain insufficiently reliable, as evidenced by significant discrepancies between their outcomes and authentic human actions. Our investigation begins with a systematic review of LLM-based human simulations in social, economic, policy, and psychological contexts, identify

Read source article
arXiv cs.CL (NLP)Research

Evolutionary Guided Decoding: Iterative Value Refinement for LLMs

arXiv:2503.02368v4 Announce Type: replace Abstract: While guided decoding, especially value-guided methods, has emerged as a cost-effective alternative for controlling language model outputs without re-training models, its effectiveness is limited by the accuracy of the value function. We identify that this inaccuracy stems from a core distributional gap: existing methods train static value functions on trajectories sampled exclusively from the base policy, which inherently confines their traini

Read source article
arXiv cs.CL (NLP)Research

Leveraging Natural Language Processing to Unravel the Mystery of Life: A Review of NLP Approaches in Genomics, Transcriptomics, and Proteomics

arXiv:2506.02212v2 Announce Type: replace Abstract: Natural Language Processing (NLP) has transformed various fields beyond linguistics by applying techniques originally developed for human language to the analysis of biological sequences. This review explores the application of NLP methods to biological sequence data, focusing on genomics, transcriptomics, and proteomics. We examine how various NLP methods, from classic approaches like word2vec to advanced models employing transformers and hyen

Read source article
arXiv cs.CL (NLP)Research

Curriculum-Guided Layer Scaling for Language Model Pretraining

arXiv:2506.11389v4 Announce Type: replace Abstract: As the cost of pretraining large language models grows, there is continued interest in strategies to improve learning efficiency during this core training stage. Motivated by cognitive development, where humans gradually build knowledge as their brains mature, we propose Curriculum-Guided Layer Scaling (CGLS), a framework for compute-efficient pretraining that synchronizes increasing data difficulty with model growth through progressive layer s

Read source article
The Guardian AIBusiness

Indecent proposal: why social media’s rebrand of surveillance tech normalises harassment and non-consensual filming | Maggie Zhou

<p>By selling AI glasses as aspirational, cool and fashion-forward, tech elites are trying to pacify their entry into the mainstream world</p><p>We have a habit of dismissing social media trends as inane and vapid while ignoring the disturbing undercurrent. A few weeks ago I was reminded of that when I saw an Instagram carousel by British fashion personality Alexa Chung. Shared with her 6 million followers, she showed different outfits through screenshots of herself entering and leaving her home

Read source article
Indecent proposal: why social media’s rebrand of surveillance tech normalises harassment and non-consensual filming | Maggie Zhou
Hacker News AILLMs

Panoptes – AI audit and alignment layer

Article URL: https://github.com/miggy-code/Panoptes Comments URL: https://news.ycombinator.com/item?id=48813183 Points: 1 # Comments: 1

Read source article
Hacker News: Show HN

Show HN: Emem.dev – signed earth memory for physical AI

Hi everyone, We added 3D to emem.dev , now you can get the signed answers as well as constructed Gaussian splats all directly from https://emem.dev for free. Give us a star at - https://github.com/Vortx-AI/emem Comments URL: https://news.ycombinator.com/item?id=48813148 Points: 1 # Comments: 0

Read source article
Hacker News: Show HN

Show HN: Storytelling for coding agents, using Pixar's story process

I taught my coding agent to write stories the way Pixar makes films. It wrote 2 illustrated books. Repo is MIT; ideas from Karina Nguyen's Lux Summit talk. Comments URL: https://news.ycombinator.com/item?id=48813068 Points: 2 # Comments: 0

Read source article
Hacker News AILLMs

OSS Local AI Workspace

Article URL: https://www.usestitch.ai/ Comments URL: https://news.ycombinator.com/item?id=48813043 Points: 4 # Comments: 2

Read source article
Product HuntTools

Kickbacks CLI

<p> The terminal and Mac menu bar companion for Kickbacks.ai </p> <p> <a href="https://www.producthunt.com/products/kickbacks-cli?utm_campaign=producthunt-atom-posts-feed&utm_medium=rss-feed&utm_source=producthunt-atom-posts-feed">Discussion</a> | <a href="https://www.producthunt.com/r/p/1189818?app_id=339">Link</a> </p>

Read source article
Hacker News AILLMs

Millfolio – my take on local/hybrid AI

Article URL: https://millfolio.app/blog/send-the-program-to-your-data/ Comments URL: https://news.ycombinator.com/item?id=48812775 Points: 1 # Comments: 0

Read source article
Hacker News AILLMs

How Much Is AI Manipulating Us?

Article URL: https://americanrefugees.substack.com/p/how-much-is-ai-manipulating-us Comments URL: https://news.ycombinator.com/item?id=48812760 Points: 4 # Comments: 8

Read source article
Hacker News AILLMs

The AI-Native Founder

Article URL: https://www.wensenwu.com/thoughts/ai-native-founder Comments URL: https://news.ycombinator.com/item?id=48812374 Points: 3 # Comments: 0

Read source article
Hacker News Ask

The Hard Parts of Streaming Audio in Voice Agents

Blog: https://gokuljs.com/blogs/when-latency-becomes-audible code: https://github.com/gokuljs/GoSFU Comments URL: https://news.ycombinator.com/item?id=48812258 Points: 2 # Comments: 1

Read source article
Hacker News LLMLLMs

Jackrong LLM Fine-Tuning Guide

Article URL: https://github.com/R6410418/Jackrong-llm-finetuning-guide Comments URL: https://news.ycombinator.com/item?id=48812171 Points: 2 # Comments: 0

Read source article