AiAnyTool - Best AI Tools Directory and Artificial Intelligence Software Hub Logo
Loading theme toggle

Real-Time Coverage

AI News Today

Live

24855 stories from 30+ sources, refreshed continuously.

arXiv cs.CL (NLP)Research

An MLIR-Based Compilation Method for Large Language Models

arXiv cs.CL (NLP)Research

How Much Human Label Variation Does Formal Semantic Structure Explain?: Group-Level Effects and Item-Level Ceilings in NLI

arXiv cs.CL (NLP)Research

Induction in Both Directions: A Mechanistic Analysis of In-Context Learning in Masked Diffusion Language Models

arXiv cs.CL (NLP)Research

From Plausible to Actionable: A Position on LLM Self-Explanations

arXiv cs.CL (NLP)Research

BayesPO: Bayesian Prompt Optimization via Parallel-Tempered Gradient-Guided Discrete MCMC

arXiv cs.CL (NLP)Research

Loop the Loopies!

arXiv cs.CL (NLP)Research

Frontier AI performance across the business disciplines: a case-grounded benchmark of knowledge work and analytical reasoning

arXiv cs.CL (NLP)Research

Frontier Language Models Struggle to Copy: Text Can Be Better Viewed in 2D

arXiv cs.CL (NLP)Research

Controlling Implicit Shortcut Reliance in L2 Spoken English Auto-markers

arXiv cs.CL (NLP)Research

CoWeaver: A Bi-directional, Learnable and Explainable Matching Engine for Mixed Human-Agent Science Collaboration

arXiv cs.CL (NLP)Research

AI Watermark Evidence Fails Forensic Readiness: An Empirical Evaluation

arXiv cs.AIResearch

GraphDx: A Cost-Aware Knowledge-Enhanced Multi-Agent Framework for Sequential Diagnosis

arXiv cs.AIResearch

Causal-Audit: Explicit and Auditable Graph-based Reasoning via Target-Aware Causal Chain Construction

arXiv cs.AIResearch

Cura 1T: Specialized Model for Agentic Healthcare

arXiv cs.AIResearch

AnovaX: A Local, Multi-Agent Voice Assistant with LLM Planning, Typed Executors, and Adaptive Recovery

arXiv cs.AIResearch

Precise but Uncoupled: Reviewer Precision Does Not Guarantee Critique Uptake in Multi-Agent Math Reasoning

arXiv cs.AIResearch

DrawingVQA: A Real-World Benchmark for Multi-Depth Visual-Textual Reasoning on Construction Drawings

arXiv cs.AIResearch

Do Coding Agents Need Executable World Models, Simplification, and Verification to Solve ARC-AGI-3?

arXiv cs.AIResearch

A Critical Analysis of Trustworthy AI Tools, Mark Frameworks, and the Implementation Chasms

arXiv cs.AIResearch

Logic, Optimization, and Artificial Intelligence

arXiv cs.AIResearch

SeerGuard: A Safety Framework for Mobile GUI Agents via World Model Prediction

arXiv cs.AIResearch

MGDT: MLLM-Guided Diffusion Transformer with Relation-Adaptive Mixture-of-Experts for Multimodal Knowledge Graph Completion

arXiv cs.AIResearch

Neuro-Symbolic AI for LEED compliance: Document-Centric Benchmarking, Deterministic Numeric Checking, and When Multimodal Hurts

arXiv cs.AIResearch

ToolVerse: Unlocking Massive Environments and Long-Horizon Tasks for Agentic Reinforcement Learning

arXiv cs.AIResearch

S1-Omni: A Unified Multimodal Reasoning Model for Scientific Understanding, Prediction, and Generation

arXiv cs.AIResearch

Behavioral Controllability of Agentic Models for Information Extraction: From Fixed Workflows to Reflective Agents

arXiv cs.AIResearch

NeurOWL: An LLM-Based Neural-symbolic Framework for Incomplete OWL Ontology Reasoning

arXiv cs.AIResearch

AgentFAIR: A Multi-Agent Collaborative Framework for FAIRness Evaluation of Geospatial Datasets

arXiv cs.AIResearch

Knowledge-Centric Agents for Workflow Generation

arXiv cs.AIResearch

DSWorld: A Data Science World Model for Efficient Autonomous Agents

arXiv cs.AIResearch

A Formally Grounded ODRL Evaluator: Implementation and Comparison

arXiv cs.CVResearch

DiTango: Cost-Effective Parallel Diffusion Generation with Selective Attention State Reuse

arXiv cs.CVResearch

Model Merging for Medical LVLMs: A Benchmark and a Winner-Take-All Approach

arXiv cs.CVResearch

PE-Field 4D: Video Generation Models as Canvas

arXiv cs.CVResearch

Efficient Frame Selection for Long Videos at Test Time with Attention-Based MLLM Selectors

arXiv cs.CVResearch

Per-Stroke Temporal Control for Text-to-Motion via Action Units and Action-Detection Guidance

arXiv cs.CVResearch

Event3R: Asynchronous-to-Global 3D Reconstruction from Event Camera via Spatial-Temporal Feature Aggregation

arXiv cs.CVResearch

IoUPD: IoU-Aware Privileged Distillation for Visual Grounding with Multimodal Large Language Models

arXiv cs.CVResearch

Personalized Image Aesthetic Assessment via Preference-rich Sample Mining and Cohort Merging

arXiv cs.CVResearch

Exo2EgoPose: Leveraging Exocentric Demonstrations for Vision-Language guided Egocentric 3D Hand Pose Forecasting