AiAnyTool - Best AI Tools Directory and Artificial Intelligence Software Hub Logo
Loading theme toggle

Real-Time Coverage

AI News Today

Live

9394 stories from 30+ sources, refreshed continuously.

r/MachineLearningResearch

My OCR model mislabels section titles as body text. Is a CRF the right fix, or am I overcomplicating it? [P]

MarkTechPostResearch

NVIDIA Releases Cosmos 3 Edge: A 4B-Parameter Open World Model That Reasons and Generates Robot Actions On-Device

r/MachineLearningResearch

Reproducing OpenAI’s “persistently beneficial models” - GRPO trait install barely moves. Ideas? [P] [R]

arXiv cs.AIResearch

Design and Validation of a Lightweight 1D CNN for Affective Touch Classification in Soft Plush Companions

arXiv cs.AIResearch

Some Large Language Models Exhibit Consistent Risk Attitudes

arXiv cs.AIResearch

A Survey on GNN-based Link Prediction: Techniques, Applications, and Challenges

arXiv cs.AIResearch

PlanFlip: Attacking Multi-Agent LLM Systems via Planning-Phase Prompt Injection

arXiv cs.AIResearch

Deterministic Replay for AI Agent Systems

arXiv cs.AIResearch

Generative Ontology Induction: Domain-Agnostic Schema Discovery from Document Corpora Using Large Language Models

arXiv cs.AIResearch

Democratizing AI with Small Language Models: Structured Benchmarking and Parameter-Efficient Fine-Tuning for Local Deployment

arXiv cs.AIResearch

Masked Diffusion Language Models are Strong and Steerable Text-Based World Models for Agentic RL

arXiv cs.AIResearch

It Takes 8 Tokens: Weak-to-Strong Off-Policy RL via Auxiliary Branches

arXiv cs.AIResearch

PPO-HSC: An Exploratory Reinforcement Learning Framework Based on Wide-Area Policy Coverage Optimization

arXiv cs.AIResearch

JUMP: Single-Pass Membership Inference on Fine-Tuned Diffusion Language Models

arXiv cs.AIResearch

A Survey on the Verification of Reinforcement Learning Policies

arXiv cs.AIResearch

Accurate and Efficient Long-Term Memory for LLM Agents

arXiv cs.AIResearch

Symbolic Augmentation Closes a Canonical-Equivalence Blind Spot in Neural Fact-Checkers

arXiv cs.AIResearch

SelKV: Selective KV Cache Merging with Per-Token Merge-or-Drop and Attention Compensation

arXiv cs.AIResearch

RAIL Guard: Closing the Evaluation-to-Remediation Gap in Responsible AI for LLM Agents

arXiv cs.AIResearch

Generalist AI Control: Towards Multi-purpose Adaptive Algorithms

arXiv cs.AIResearch

LaCache: Exact Caching and Precision-Adaptive Inference for Diffusion Large Language Models

arXiv cs.AIResearch

When to Plan: Learning to Select Between Reactive Control and Deliberative Planning

arXiv cs.AIResearch

SEER: Supervised Learning to Control Energetic Reasoning

arXiv cs.CVResearch

Why do CNNs excel at feature extraction? A mathematical explanation

arXiv cs.CVResearch

3D Motion Perception of Binocular Vision Target with PID-CNN

arXiv cs.LGResearch

Reinforcement Learning-Guided NSGA-II Enhanced with Gray Relational Coefficient for Multi-Objective Optimization: Application to NASDAQ Portfolio Optimization

arXiv cs.LGResearch

LLM Unlearning for Cyber Defense: A Survey on Methods, Challenges, and Emerging Threats

arXiv cs.LGResearch

Operator-Aware Mixed-Precision Tolerance Calibration for Tensor Kernels

arXiv cs.LGResearch

Orthogonal Gradient Constraints Shape Noisy-Label Memorization Dynamics

arXiv cs.LGResearch

BACON: Budgeted Human Calibration for Modeling and Evaluation with Multiple AI Judges

arXiv cs.LGResearch

Self-Evolving Just-In-Time Memory for Proactive Embodied Safety

arXiv cs.LGResearch

Learning Spatio-Temporal Foundation Models from Pure Synthetic Data

arXiv cs.LGResearch

Quantifying Ranking Uncertainty in LLM Benchmarks

arXiv cs.LGResearch

PsiLogic: Chaos-Aware Active Cancellation for Adam with a Fair Cross-Domain Benchmark

arXiv cs.LGResearch

Scaling Limits of Constant-Stepsize SGD at Flat Minima

arXiv cs.LGResearch

EA-RMENet -- Path Loss Prediction in Urban Environments using Deep Learning

arXiv cs.LGResearch

Compact convolutional neural networks for AI-based drone detection system

arXiv cs.LGResearch

Leakage-Robust Evaluation and Data-Scale Sensitivity of Attention-Enhanced Multi-Task Learning for Joint Fault Diagnosis and Remaining Useful Life Estimation

arXiv cs.LGResearch

Feedback Attribution and Representation Geometry: Metrics for Comparing Individual and Shared Rewards in MARL

arXiv cs.LGResearch

Hierarchical Domain Generalization