Skip to content
HeyJared

Warp

The open platform for automating development. Infrastructure to build, measure, and interact with agents across the SDLC warp.dev/factories

YouTube · US · Official site

Indexed videos, last 90 days
16
Latest publication
Sep 28, 2026
Audience
~113K subscribers
Earliest in this view
Jul 7, 2026

Latest videos

  1. Video · Sep 28, 2026

    Warp Factories - the complete guide (opens the original)

    Excerpt · Positive tone · 19 min · 1,947 views by Sep 30, 2026

    Read excerpt

    A complete tour of Warp Factories, infrastructure to build your own AI software factory. I'll show how to set up agents and automations, launch work from Slack, then use spending dashboards, scoring, self-improvement, and benchmarks to improve results. 0:00 Overview 0:55 Create your first factory 3:22 Manage agent orchestration 4:42 Connect agents to automations 6:25 Set up integrations, MCPs, and secrets 8:09 Start a factory run from Slack 10:48 View the metrics dashboard 13:18 Using scoring ag

  2. Video · Sep 17, 2026

    You should have an agent that grades your agents (opens the original)

    Excerpt · 6 min · 1,509 views by Sep 30, 2026

    Read excerpt

    Try Warp Factories: warp.dev/factories/request-access Let's measure the quality of coding agents (Claude Code, Codex, etc) with scoring agents. These use LLM-as-a-judge to measure output quality along metrics you care about, like code quality, efficiency, verbosity, and task compliance. These feed self-improvement and benchmarking workflows to improve your software factory automatically.

  3. Video · Sep 4, 2026

    The secret to cheaper, smarter coding agents (opens the original)

    Excerpt · Neutral tone · 4 min · 1,619 views by Sep 30, 2026

    Read excerpt

    With Warp Factories Benchmarking it’s now possible to build custom coding agent benchmarks (Claude, Codex, Grok, Kimi, etc) on your team’s real data and workflows with a few clicks. This is a deep dive of Factory Benchmarks: how to read the report, pick the right models, and build custom routers to reduce cost-per-PR.

  4. Video · Sep 3, 2026

    How to build a model bench on your own coding tasks (opens the original)

    Excerpt · Positive tone · 6 min · 2,524 views by Sep 30, 2026

    Read excerpt

    Factory Benchmarks are the model bench generated from your own agent coding tasks. They help you measure, test and improve coding agents by replaying past agent runs. We used these benchmarks to cut our cost-per-PR by 63% Get early access: warp.dev/factories/request-access

  5. Video · Aug 31, 2026

    Your agent skills should improve themselves (opens the original)

    Excerpt · Critical tone · 3 min · 4,306 views by Sep 30, 2026

    Read excerpt

    What if coding agents (Claude Code, Codex, Warp, etc) could improve Skills by reviewing past conversations? Here's the three step loop: - Score conversations from criteria you define - Isolate failures - Generate skill improvements 0:00 Why agents need self-improvement loops 0:25 Three-step scoring system 1:12 Configuring a code quality scorer 2:14 Self-improvement agents 3:13 Closing the loop

Publishing over time

Last 90 days. Choose a month to open its work.

Recurring subjects

Named in the text we hold. One piece can cover several.

Audience

~113K subscribers

Measured Sep 19, 2026

Source's subscribers, not the number who saw an individual piece.

How this was measured

About this data

Counts cover the work we have indexed. Tone needs enough text and a confident classification. Excerpts and episode notes are not full articles or transcripts.

Identity or attribution wrong? Suggest a correction.

See coverage about Warp