Search

Tag: #research297 results

Inference Scaling vs Larger Tasks Clarified

Inference Scaling vs Larger Tasks Clarified

Distinguishes rising LLM inference compute into larger tasks (human-like linear scaling) vs. true inefficiency beyond human cost fractions. Uses Pareto frontier of budget vs. 50% reliability time-horizon. Argues much progress is bigger tasks, not unsustainable scaling.

AI Alignment ForumCommunityFeb 11#research#inference-scaling#none
JiTTesting Revives Testing in Agentic Era

JiTTesting Revives Testing in Agentic Era

Agentic development accelerates code writing, review, and shipping, outpacing traditional testing. Meta introduces JiTTesting to enable faster, just-in-time bug detection as code lands. This evolves testing frameworks for modern development speeds.

Meta Engineering BlogOfficialFeb 11#research#meta#jit-testing
Harness Engineering with Codex

Harness Engineering with Codex

OpenAI blog post explores harness engineering using Codex in an agent-first world. Authored by Ryan Lopopolo, Member of the Technical Staff. Focuses on leveraging Codex for advanced AI agent development.

OpenAI BlogOfficialFeb 11#research#openai#codex
Page 30 of 30