Search

Tag: #research297 results

AI Advances via Compute, Not Smarts

AI Advances via Compute, Not Smarts

MIT report shows frontier models like OpenAI's GPT rely on more computing power rather than smarter algorithms. This scaling approach drives progress but hikes costs. The trend raises questions on sustainability.

ZDNet AIMediaFeb 13#research#mit#openai
Hyperparameter Transfer Across All Scaling Axes

Hyperparameter Transfer Across All Scaling Axes

Apple ML extends μP for hyperparameter transfer across model sizes, modules, width, depth, batch, and duration. Introduces Complete(d) Parameterisation unifying width-depth scaling. Enables optimal base hyperparameters search at small scales for large model transfer.

Apple Machine LearningOfficialFeb 13#research#apple-ml#mu-p
Hyperparam Transfer Across All Scales

Hyperparam Transfer Across All Scales

Apple extends μP for hyperparameter transfer across modules, width, depth, batch, and duration. Introduces Complete(d) Parameterisation unifying width-depth scaling. Enables optimal hypers from small to large models.

Apple Machine LearningOfficialFeb 13#research#apple-ml#mu-p
Faster Rates for Federated VIs

Faster Rates for Federated VIs

Apple advances federated optimization for stochastic variational inequalities. Establishes improved convergence rates closing gap with convex optimization. Refined analysis boosts Local Extra SGD for smooth monotone VIs.

Apple Machine LearningOfficialFeb 13#research#apple-ml#local-extra-sgd
Page 13 of 30