Search

Tag: #generalization12 results

Rethinking Reasoning SFT Generalization

Rethinking Reasoning SFT Generalization

Challenges claim that SFT memorizes while RL generalizes in reasoning. Cross-domain generalization depends on optimization dynamics, data quality, and base model capability. Reveals dip-and-recovery training pattern, asymmetric effects on reasoning vs. safety.

ArXiv AIResearchApr 9#generalization#chain-of-thought
RL Environments: Pixels to Semantic Agents

RL Environments: Pixels to Semantic Agents

This empirical study analyzes over 2,000 RL publications to map the evolution from physical simulations to LLM-driven agents. It proposes a multi-dimensional taxonomy revealing a bifurcation into 'Semantic Prior' (LLM-dominated) and 'Domain-Specific Generalization' ecosystems. The work uncovers cognitive fingerprints and offers a roadmap for next-gen embodied simulators.

ArXiv AIResearchMar 27#environments#taxonomy#embodied-ai
AI Inner Speech Speeds Up Learning

AI Inner Speech Speeds Up Learning

OIST researchers demonstrate that AI models with 'self-mumbling' internal dialogue and working memory achieve superior learning efficiency, multitask generalization, and adaptation on sparse data. Mimicking human inner speech, the method excels in complex tasks like sequence reversal. Findings published in Neural Computation.

Poggio: AI Needs Maxwell-Like Theory

Poggio: AI Needs Maxwell-Like Theory

MIT's Tomaso Poggio likens current AI to pre-Maxwell electricity, driven by engineering without deep theory. He advocates sparse compositionality—combining simple low-dim functions—for generalization in intelligent systems. Discussion covers his learning theories from kernel machines to deep nets.

Page 1 of 2