Ysra SWE
The software engineer.
Investigates real repositories, makes changes across files, runs the checks, and hands back work a person can review.
Coming soon · in training
Our first software-engineering model — trained on how work actually gets finished.
Frontier models already know how to code. Ysra SWE is being trained on something rarer: real engineering work that was run, checked and proven — so it learns to finish the job, not just write the code.
Why build our own model
The data behind it
Ysra has already done a great deal of real engineering. Very little of it is imitated as-is — only work that was checked and proven. The rest becomes tests, lessons and harder tasks.
Sessions, outcomes and traces from real Ysra work
Whole sessions with a known outcome
Two attempts at one task, one clearly better
Where “done” turned out not to be
Real repository, exact starting commit, hidden checks
Work that passed every check — the only runs we imitate
A set of real tasks, kept apart. Never trained on, never tuned against — kept aside so the release score means something.
How work is scored
A model trained to make tests pass will learn to make tests pass — even by deleting them. So Ysra SWE is rewarded by checks it can’t see, and gets nothing for a shortcut.
The kind of check we use
Task: rename calc_total to compute_total, keeping the old name working.
def compute_total(items): …
def calc_total(items):
return compute_total(items)What earns reward
Break the core behaviour and nothing else counts. A neat diff can’t rescue broken work.
How it learns
First it learns from work that already held up. Then it practises on fresh, real tasks — and only attempts that pass hidden checks shape the next version.
What training is aiming for
More tasks solved, for less. Whether you run it lean or let it dig deep, Ysra SWE is built to get more real engineering done at every level of effort.
How we’ll judge it
Ysra SWE will be measured against its own starting point on a held-out set of real engineering tasks it has never seen, through the same Ysra agent you use.
If any of those fail, it doesn’t ship. We’ll publish the results with the release.
The Ysra model family
Ysra SWE comes first. Every model after it has to earn its place with verified results.
Ysra SWE
Investigates real repositories, makes changes across files, runs the checks, and hands back work a person can review.
Ysra Fast
A smaller model for everyday work that hands harder tasks up to Ysra SWE.
Ysra Arabic & Darija
Arabic, Moroccan Darija, French and English — including how people mix them at work.
Ysra Finance
Invoices, reconciliation, ledgers and VAT, where answers can be verified exactly.
Ysra Reviewer
A focused critic that spots gaps, risks and missing evidence before work comes back to you.
Plain facts
No borrowed benchmarks and no “trained from scratch”. Just what it is, where it comes from, and how your data is treated.
Coming soon
Early access goes to teams already handing real engineering work to Ysra. Tell us about yours.