Scores that understand your code.

Scores that understand your code.

Scores that understand your code.

Silk 1 is Weave’s new code output model. 3.3x more accurate, with plain-language reasoning for every score and calibration that learns from your team. Same scale. Same API. Better signal.

Silk 1 is Weave’s new code output model. 3.3x more accurate, with plain-language reasoning for every score and calibration that learns from your team. Same scale. Same API. Better signal.

Human vs. AI output

last 90 days

Proportion of code output generated by AI compared to all code output

800

600

400

200

0

2026-1-5

2026-1-19

2026-2-2

2026-2-16

2026-3-2

2026-3-16

2026-3-30

Human output

AI output

Old model

Old model

±3.3h

Silk 1

Silk 1

±1.0h

Average error vs expert estimate

3.3×

lower average prediction error · 88% of PRs land within ±1h

Who it’s for

Built for every layer of the engineering org

Built for every layer of the engineering org

Silk 1 surfaces differently depending on where you sit.

VPs and CTOs

board-ready rollups

board-ready rollups

board-ready rollups

Engineering managers

output you can share

output you can share

output you can share

Engineers

score plus the reasoning

score plus the reasoning

score plus the reasoning

Silk 1

one scale · one API · one report format

one scale · one API · one report format

Where Silk 1 shows up

Where Silk 1 shows up

one model, three views

one model, three views

Engineering managers

Accurate output data across your team without second-guessing the numbers. Automated reporting you can trust and share with leadership as-is.

VPs and CTOs

Board-ready data on engineering productivity powered by a model that understands your codebase. Make investment decisions based on signal, not estimates.

Engineers

Transparent scores with clear reasoning. Understand how your work is measured and flag scores that don’t look right. Your overrides make the model better for everyone.

What’s better

Silk 1 reads the full diff and reasons about intent, risk, and interdependencies across files.

3.3x

3.3x

Reduction in average prediction error

88%

Of PRs scored within ±1h of expert estimate

35x

More model capacity than the original

Old model

Old model

Sees 3 lines changed. Scores for line count.

Scored complexity

3 lines

Silk 1

Silk 1

Silk 1

Understands the downstream impact across 40 files. Scores for true complexity.

Understands the downstream impact across 40 files. Scores for true complexity.

Scored complexity

Scored complexity

40 files

40 files

Example

Example

A 3-line config change that unlocks a migration path across 40 files.

config/migrate.yml

strategy: legacy

+

strategy: incremental

+

parallel: true

What the model does

What the model does

What the model does

Four capabilities the old model didn’t have.

Four capabilities the old model didn’t have.

Four capabilities the old model didn’t have.

Deeper code comprehension

Deeper code comprehension

Understands what a code change does and why it’s complex. Analyzes the complexity of each line changed and reasons about cross-file dependencies.

Understands what a code change does and why it’s complex. Analyzes the complexity of each line changed and reasons about cross-file dependencies.

Transparent reasoning

Transparent reasoning

Every score ships with a plain-language explanation of how the model arrived at its estimate. When a score looks off, read the reasoning instead of guessing. No other engineering analytics tool offers this level of transparency.

Every score ships with a plain-language explanation of how the model arrived at its estimate. When a score looks off, read the reasoning instead of guessing. No other engineering analytics tool offers this level of transparency.

Organization-aware calibration

Organization-aware calibration

Uses reinforcement learning to learn from your team’s score overrides. The model adapts to your codebase’s conventions and complexity profile over time. The longer you use it, the more accurate it gets.

Uses reinforcement learning to learn from your team’s score overrides. The model adapts to your codebase’s conventions and complexity profile over time. The longer you use it, the more accurate it gets.

Cross-file understanding

Cross-file understanding

Reasons about how changes in one file affect others. Complex refactors and cross-cutting changes finally get the scores they deserve.

Reasons about how changes in one file affect others. Complex refactors and cross-cutting changes finally get the scores they deserve.

What stays the same

Zero migration required.

Silk 1 uses the same scoring scale, the same API, and the same report format. A “3.0” still means three hours of expert engineer effort.

Same scoring scale

Same scoring scale

Same API

Same API

Same report format

Same report format

Dashboards, integrations, and alerts work without modification

Dashboards, integrations, and alerts work without modification

Measure engineering output with the most accurate scoring model in the industry.

Get started in 5 minutes or book a demo with our team.