πŸ€–Stalecollected in 2h

Weber Optimizer Powers Autonomous ML Fork

PostLinkedIn
πŸ€–Read original on Reddit r/MachineLearning
#optimizer#hardware-entropy#ai-agents#multi-gpudeepbluedynamics/autoresearchkarpathy/autoresearchclaudegpt-4ogeminirtlsdr

πŸ’‘Physics-based optimizer + hardware entropy for autonomous MLβ€”test if it beats AdamW

⚑ 30-Second TL;DR

What Changed

Weber optimizer uses 19th-century electrodynamics bracket for per-parameter learning rate modulation based on velocity and acceleration.

Why It Matters

This fork could accelerate autonomous ML research by introducing novel optimizers and true randomness, potentially stabilizing training and improving results. Community-driven improvements make it accessible for experimentation on various hardware.

What To Do Next

Clone DeepBlueDynamics/autoresearch and benchmark Weber optimizer against AdamW on your H100 setup.

Who should care:Researchers & Academics

Key Points

  • β€’Weber optimizer uses 19th-century electrodynamics bracket for per-parameter learning rate modulation based on velocity and acceleration.
  • β€’True hardware random seeding via RTL-SDR radio receiver capturing ADC noise.
  • β€’Multi-provider agent.py harness supports Claude, GPT-4o, Gemini with 10 tools and thermodynamic memory.
  • β€’Multi-GPU support for H100 and consumer GPUs, Docker container included.
  • β€’Baseline improvement to 0.9697 val/bpb from community experiments.
πŸ“°

Weekly AI Recap

Read this week's curated digest of top AI events β†’

πŸ‘‰Related Updates

AI-curated news aggregator. All content rights belong to original publishers.
Original source: Reddit r/MachineLearning β†—

This is a summary, not the original. Read the source, or get the weekly briefing.

Weekly AI briefing

One email a week. Unsubscribe anytime.