雷达日报 · 2026-09-10
今日 65 条信号 · SI 正常 · 精选 4
🎯 今日精选
1. Procedural Graphs: Self-Evolving Execution Structures for LLM Agents
2609.09153 · ? · HF 🔥23 · #untagged
核心 idea:Large language models are increasingly deployed as agents that plan over long horizons and act through external tools.
arXiv
2. DianShi-RxnDB: A Large-Scale, Fine-Grained Organic Reaction Data Platform Built via a Fully Automated Pipeline for Researchers and AI Agents
2609.06703 · ? · HF 🔥4 · #untagged
核心 idea:High-quality structured organic reaction data are essential for developing artificial intelligence for chemistry (AI4Ch…
arXiv
3. EVOHARNESSBENCH: Can Your Agents Keep Pace with an Evolving Harness?
2609.04280 · ? · HF 🔥4 · #untagged
核心 idea:Modern LLM-based agents operate through a harness of tools, reusable skills, and specialist agents that shapes what the…
arXiv
4. Scores Alone Do Not Prove Discovery: The Discovery Certification Protocol for Auditing AI Research Agents
2609.09219 · ? · HF 🔥11 · #untagged
核心 idea:AI research agents combine prior knowledge, public sources, and experimental feedback to produce useful results.
arXiv
📋 速览(其余 61 条)
- SAEScientist-Bench: Can AI Agents Conduct Autonomous SAE Interpretability Research?
2609.09113· 🔥7 - Omni Interaction Agent Technical Report
2609.08977· 🔥121 - Show-Harness: Just a VLM Agent Can Play Robots
2609.10522· 🔥31 - Programmable World Model
2609.10540· 🔥21 - RoboSPA: Can VLA Models Go Beyond Simple Scenes and Short-Horizon Tasks?
2609.05324· 🔥21 - WearableQA: A Benchmark for Health Reasoning over Real-World Wearable Data
2609.05405· 🔥16 - Counter-Swarm Doctrine: Containing Coordinated Agent Intrusions
2609.06140· 🔥3 - AgenticGen: Reward-Guided Agentic Video Generation for Advertising
2609.09187· 🔥2 - MasterControl Seventeen Every Time
2609.03209· 🔥1 - Stencil Computation at the Intersection of AI and HPC
2609.10368· SI 83 - OpenWAM: An Open, Modular Exploration Towards Systematic World-Action Model Pretraining
2609.07398· 🔥55 - Steering Geometry: Validating Human Value Geometry in LLM Steering Space
2609.06289· 🔥27 - VDiff-Bench: A Challenging Benchmark for Fine-Grained Image Difference Identification
2609.06245· 🔥22 - Harnessing CLIP and DINO: An Uncertainty-Aware Cascaded Fusion Network for Generalizable Deepfake Image Detection
2609.07670· 🔥15 - A*-Thought-V2: Efficient Latent Reasoning via Geometric Dynamics of LLM
2609.07821· 🔥11 - Φ-Bench: Can Large Language Models Engineer the Infrastructure That Powers Them?
2609.10226· 🔥3 - Co-Evolving Harnesses and Models: On-Policy Correction Helps Weaker Models Catch Up Where Imitation Fails
2609.09134· 🔥1 - RESCUE-BENCH: Towards Relation-Aware Multi-Party Emotional Support Conversation Systems
2609.09657 - AXON: A ROS 2 RMW with Shared-Memory/QUIC Transport and QKD/ML-KEM Key Establishment
2609.10024· SI 65 - Towards a Quantum Erlangen Program
2609.09167· SI 59 - Academia x Industry: The Role of Fundamentals for Silicon in an AI Native Era
2609.09344· SI 46 - Measuring Language Transfer in Robot Policies: Adding Greek to a Cosmos3 Vision-Language-Action Policy
2609.07470· 🔥18 - ReactVAU: A Slow-Fast Decoupled Framework for Streaming Video Anomaly Understanding
2609.07941· 🔥15 - Recognition-Refusal Misalignment in LLMs: Why Models Answer Structurally Unanswerable Questions
2608.29109· 🔥15 - Revisiting Complete Reasoning Traces for Post-Training
2609.07103· 🔥5 - Difficulty-Adaptive Tree-Structured Policy Optimization for Expanding Reasoning Coverage in RLVR
2609.08650 - Where Does the Human End? Creative Agency with Generative AI across Five Years of Chinese Digital Painting
2609.09333· SI 74 - A Multi-Model Non-Intrusive Reduced-Order Framework for Parametric Erosion Prediction via Kinematic Cross-Moment Compression
2609.09997· SI 62 - Violet: Enabling Full Virtualization for M-mode RTOS on RISC-V
2609.09833· SI 51 - An $O(1/T^3)$ algorithm for minimizing convex quadratic functions over the $L_1$ ball
2609.10314· SI 50 - AuK Technical Report: An Open-Source Foundational Model for Speech Generation and Editing
2609.08936· 🔥182 - BeaconKV: Key-Value Cache Compression Guided by Beacon Queries for Efficient Large Reasoning Model Inference
2609.04971· 🔥32 - Reason Through the Latent! Making Latent Visual Reasoning Necessary
2609.06746· 🔥30 - CosmoH2G: A Hand-to-Gripper Transfer Dataset and Baseline Method for Object Manipulation with Complex Spatial Movements
2609.07498· 🔥25 - CoVeR: Coverage-Based Token Pruning for Multi-View 3D Reasoning in VLMs
2609.08345· 🔥20 - TransNormal-2: Geometry-Grounded Rectified Flow with Edge-Aware Decoding for Precise Normal Estimation
2609.06665· 🔥18 - Cadence: Error-Bounded Lossy Compression of Demand Time Series with a Time-Series Foundation Model
2609.06008· 🔥16 - SynthGait-19K: A Physically Grounded Synthetic Video Dataset for Gait Parameter Estimation
2609.08108· 🔥15 - StudyBench: Can Self-Evolution Squeeze Textbooks for Olympiad Capability?
2609.00787· 🔥12 - Diffs vs. Whole Files: An Empirical Comparison of Iterative Edit-Based and Direct Generation for Flutter/Dart Code Models
2609.05779· 🔥1 - Puppeteer: Object-Grounded Posture-Aware Co-Speech Gesture Generation
2609.00369· 🔥1 - Planetary Nebula Central Stars as Tracers of Planetary Nebula-Star Cluster Associations in the Galaxy
2609.09618· SI 83 - The magnetic field in M16: new results from the JCMT BISTRO survey
2609.09913· SI 78 - Gradient-Enhanced Proximal Algorithms for Mean Field Planning on Surfaces
2609.10370· SI 74 - Physics-Informed Multi-Task Surrogate Model for the Martian Nightside Thermosphere
2609.10077· SI 70 - Chance, Persistent Advantage, and the Generative-AI Era in Open-Source Package Careers
2609.09687· SI 60 - Exact-Form Regret for Gradient Descent, Mirror Descent and Follow-the-Regularized-Leader
2609.09466· SI 48 - Field-level prediction of mid-plane stress tensor fields in concrete target penetration: a cross-velocity graph neural operator surrogate
2609.10032· SI 45 - Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation
2609.08798· 🔥84 - GE-Act 2.0: Pretraining and Scaling a World-Action Model for Robotic Manipulation
2609.05588· 🔥49 - Marigold V2: Revisiting Diffusion Transformers for Monocular Depth Estimation
2609.08084· 🔥45 - Mask Forcing: Improving Autoregressive Video Diffusion Distillation via Dual-Noise Masking Rollout
2609.09123· 🔥43 - TANGO: Humanoid Navigation in Cluttered Environments with a Whole-Body Vision-Language-Action Model
2609.09158· 🔥19 - What Did I Just Say? Self-Listening for Full-Duplex Speech Models
2609.05592· 🔥16 - Encoded Early, Used Late: Where Transformers Begin to Act on an Inferred Partner's Expertise
2609.07139· 🔥15 - SQS: Bayesian DNN Compression through Sparse Quantized Sub-distributions
2510.08999· 🔥15 - RelightFormer: Feed-forward Generative Transformer for Multiview Object Relighting
2609.07414· 🔥8 - NOAH: Learning the Full Patient Journey. A Longitudinal Multimodal Time-Aware Model for Representation and Forecasting
2609.09140· 🔥3 - Graph Machine: Towards Better Pretraining via Edges
2609.02881· 🔥2 - RenderFormer-V2: Neural Rendering with Heterogeneous Scene Primitives
2609.05738· 🔥2 - Learning 3D Editing without Paired Supervision via Generative Prior Distillation
2609.04942
🔥 HF 热度榜(Top 5)
- AuK Technical Report: An Open-Source Foundational Model for Speech Generation and Editing · 🔥182
- Omni Interaction Agent Technical Report · 🔥121
- Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation · 🔥84
- OpenWAM: An Open, Modular Exploration Towards Systematic World-Action Model Pretraining · 🔥55
- GE-Act 2.0: Pretraining and Scaling a World-Action Model for Robotic Manipulation · 🔥49
📡 雷达备注
- ✓ 数据源全部正常
下期:明天 21:00(北京时间)