newsfilter.io

Josh Tobin

Showing 11 of 1 transcripts.

  1. Sequoia Capital33 min

    OpenAI’s Deep Research Team on Why Reinforcement Learning is the Future for AI Agents

    Isa Fulford, Josh Tobin, Sonya Huang, Lauren Reeder

    Launched three weeks ago, OpenAI's Deep Research is an agentic system powered by a fine-tuned O3 model that executes complex, multi-hour tasks like market analysis and medical research in 5 to 30 minutes. Utilizing reinforcement learning to optimize browsing and coding strategies, the tool distinguishes itself through a pre-research clarification flow that refines user prompts for higher-quality synthesis. As part of a broader 2025 shift toward agent-driven workflows, this technology aims to amplify knowledge workers by automating information-intensive processes previously deemed too time-consuming.