AI News

⚡ 1 minute ago
1
1
Adversarial Test-Hardening for AI-Written Code: An Instrument Autopsy and a Pre-Registered Causal Estimate of the Critic Loop (arxiv.org)
2
1
Learning Reusable Hybrid Motion Priors for Humanoid Locomotion from Motion Imitation (arxiv.org)
3
1
Calibrated Tree-Neural Fusion for Fine-Grained Vegetation Community Classification (arxiv.org)
4
1
Distributional Random Forests for Complex Survey Designs (arxiv.org)
5
1
Real2Sim2Real for Vision-Language-Action Manipulation: An AMD ROCm-Based Pipeline (arxiv.org)
6
1
Spatial Prediction of Soil Microplastics and Organic Matter Using Graph Attention Networks (arxiv.org)
7
1
MEMENTO: Memory-Guided Memetic Code-as-Policy Evolution (arxiv.org)
8
1
Nearly Tight Bounds for Cross-Learning Contextual Bandits with Graphical Feedback (arxiv.org)
9
1
Flick: Few Labels Text Classification using K-Aware Intermediate Learning in Multi-Task Low-Resource Languages (arxiv.org)
10
1
An Agentic Orchestration of Atomistic Simulations (arxiv.org)
11
1
Building AI That Works: ESnet's Pragmatic Approach to AI-Driven Operational Excellence (arxiv.org)
12
1
Do Visual Features Improve Other-Initiated Repair Detection? A Dyadic Multimodal Approach (arxiv.org)
13
1
Hierarchical Grading in Large Language Models (arxiv.org)
14
1
Charging Phase Health Indicators for Battery State-of-Health Estimation: A Systematic Comparison of CC, CV, and Combined Approaches under Cross-Battery Validation (arxiv.org)
15
1
HyCE-RAG: Hypergraph Chain-of-Evidence Retrieval-Augmented Generation for Explainable Multi-hop Question Answering (arxiv.org)
16
1
CALMRec: Causally Aligned Language Memory for Long-Horizon Recommendation (arxiv.org)
17
1
ACM: Agentic Context Management for Long Horizon Tasks (arxiv.org)
18
1
A Frozen 12B Beats Frontier Models on Verified Work: 100% Accuracy, 0 Tokens, Bit-Exact, Forever (arxiv.org)
19
1
Co-Harness: Co-Evolving Harnesses and Model Weights for LLM Agents (arxiv.org)
20
1
Evaluating and Mitigating the Misguidance Effect of Buggy Code in LLM-Generated Unit Tests (arxiv.org)
21
1
Copula Based Fusion of Clinical and Genomic Machine Learning Risk Scores for Breast Cancer Risk Stratification (arxiv.org)
22
1
TriShieldRAG: A Three-Ring Defense-in-Depth Framework Against Knowledge Corruption in Retrieval-Augmented Generation (arxiv.org)
23
1
Invariant Discovery for Networked Systems (arxiv.org)
24
1
An Empirical Study of Feature Selection Granularity (arxiv.org)
25
1
MS-GPT: Rethinking MS/MS De Novo Structure Elucidation as Spectrum-Induced Posterior Querying of a Molecule-Language Model (arxiv.org)
26
1
Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift (arxiv.org)
27
1
Where Is the Cost of Third-Party API Routers in Agentic Software Development? (arxiv.org)
28
1
Frustratingly Simple Black-Box Adaptation of Language Models via Logit Bias (arxiv.org)
29
1
Progress-conditioned Group Policy Optimization for Long-Horizon Agentic Tasks (arxiv.org)
30
1
Choosing a Text Embedding Model: A Practical Benchmarking and Decision Framework (arxiv.org)
31
1
Do Small Models Use the Law You Give Them? Context-Injected Fine-Tuning for Legal QA in Bangladesh (arxiv.org)
32
1
Physics-Informed Neural Networks for Discovering Periodic Orbits in the Gravitational Three-Body Problem (arxiv.org)
33
1
Plans Work in Mysterious Ways: Evaluating a Plan Mode for Spreadsheet Agents (arxiv.org)
34
1
Novel Claim or D\'ej\`a Vu? Rethinking "Contamination-Free'' Dynamic Evaluation for Multimodal Automated Fact-Checking (arxiv.org)
35
1
GFLAN: Generative Functional Layouts (arxiv.org)
36
1
Adaptive Data Admission and Retention for Streaming Federated Learning (arxiv.org)
37
1
A Fixed-Effects Causal Forest for Staggered Adoption, with an Application to Medicaid Expansion (arxiv.org)
38
1
Between Suppression and Collapse: Evaluating Narrative Unlearning with LENS (arxiv.org)
39
1
Quotient Tree Arithmetic: Deferred-Division Computation with Bounded Symbolic Depth and Cross-Subtree Cancellation (arxiv.org)
40
1
Variational-Ising-Attention (VIA):TailoredAttentionMattersfor Science (arxiv.org)
41
1
When Should Active RAG Retrieve? A Budget-Aware Evaluation of Utility, Calibration, and Cost (arxiv.org)
42
1
K-Survival Means (arxiv.org)
43
1
PathScale-R1: Cross-scale Reasoning for Pathological Image Analysis (arxiv.org)
44
1
A Survey of Graph Transformers: Architectures, Theories and Applications (arxiv.org)
45
1
EmotionAI: A Privacy-Preserving Computational Intelligence Pipeline for Speech-Emotion-Grounded Conversational Analysis (arxiv.org)
46
1
Probabilistic Symbolic Regression for Equation Discovery via Operator-induced and Regularized Symbolic Forests (arxiv.org)
47
1
BHARATI: Morphology-Aware Tokenizers for Classical Indian Languages with Subword Fertility Analysis (arxiv.org)
48
1
Falsifiable Commitment Planning for Self-Correcting Web Agents (arxiv.org)
49
1
xMIx: High-Performance Serving-Time Platform for Mechanistic Interpretability Apps (arxiv.org)
50
1
Semalith v1.4: A Calibrated 184M Safety Classifier Achieving State-of-the-Art Prompt-Injection Detection at 44x Fewer Parameters than Llama-Guard-3-8B (arxiv.org)