Search

Tag: #research297 results

Policy-Car Swerving Absorbs Traffic Jams

Policy-Car Swerving Absorbs Traffic Jams

Proposes SVDD-JAD strategy mimicking police swerving to suppress stop-and-go waves via slow-in/fast-out maneuvers. Analyzes five key parameters measurable with roadside detectors. SUMO simulations confirm no secondary waves triggered.

ArXiv AIResearchFeb 12#research#svdd-jad#v1
PiT-PO Boosts Equation Discovery with RL

PiT-PO Boosts Equation Discovery with RL

PiT-PO uses reinforcement learning to evolve LLMs for symbolic regression, enforcing physical validity and parsimony. It treats LLMs as adaptive generators updated by search feedback. Achieves SOTA on benchmarks and discovers novel turbulence models.

ArXiv AIResearchFeb 12#research#pit-po#v1
PELLI Framework Boosts LLM Code Quality

PELLI Framework Boosts LLM Code Quality

PELLI is an iterative framework for integrating LLMs into software generation, evaluating code on maintainability, performance, and reliability. It tests five popular LLMs across three domains using Python standards. GPT-4T and Gemini outperform others, with prompt design impacting quality.

ArXiv AIResearchFeb 12#research#pelli#v1
NSAM: Neuro-Symbolic Action Masking in DRL

NSAM: Neuro-Symbolic Action Masking in DRL

NSAM learns symbolic models and action masks automatically during DRL to avoid infeasible actions. It integrates symbolic reasoning with deep policy optimization mutually. Evaluations show improved sample efficiency and fewer violations.

ArXiv AIResearchFeb 12#research#nsam#v1
NAEs Balance Interpretability and Accuracy

NAEs Balance Interpretability and Accuracy

Neural Additive Experts use mixture-of-experts per feature with context-gated integration for flexible additivity. Targeted regularization ensures smooth transitions from additive to interactive models. Outperforms on accuracy while preserving feature explanations.

ArXiv AIResearchFeb 12#research#nae#v1
Page 20 of 30