Route reliability
At every step of an assignment there are many possible paths. Our research is making the agent take the right one, and recover fast when it doesn’t.
The Spice Lab is our research group. It doesn’t build models; it builds reliability. Wasabe runs on a leading American-made frontier model, post-trained and fine-tuned through Wasabe Spice Forge, our harness, until one agent can carry an assignment across tens of thousands of interactions.
Give a brilliant model one question and it shines. Give it a real assignment (a hundred steps, a dozen apps, tools that bite back), and every step is a chance to drift. Small errors compound; at step ten thousand, “usually right” isn’t good enough.
That is the problem the Spice Lab exists to solve. Our research isn’t about making the model smarter; the raw intelligence comes from a leading American-made model. It’s about making the agent dependable: choosing among many possible routes, choosing among many possible tools, and staying on course across tens of thousands of interactions.
Wasabe Spice Forge is our proprietary harness: the post-training, fine-tuning, and runtime strategy that turns a frontier model into a dependable coworker. It orchestrates research, tool discovery, tool use, and reasoning, and holds the agent steady across tens of thousands of interactions. It’s Wasabe’s own IP, and where most of the value is created.
Searches the web and your workspace for what the task actually needs.
Picks the right tools and apps for the job from everything available.
Calls them in order: drafting, querying, and editing across your apps.
Works over the results, plans the next move, and keeps the goal in view.
For high-stakes work, a second agent independently challenges and checks it.
Hands back a finished, sourced result you can act on.
The model supplies the raw intelligence. Everything the Spice Lab studies is what it takes to make that intelligence dependable inside real work.
At every step of an assignment there are many possible paths. Our research is making the agent take the right one, and recover fast when it doesn’t.
More than a hundred tools, each with sharp edges. We train the agent to discover, choose, and call the right one, with the right arguments, every time.
Real assignments aren’t one prompt. The Forge keeps a single agent coherent across tens of thousands of tool calls, messages, and decisions.
Spice Forge is a training strategy as much as a runtime. We post-train and fine-tune against the harness’s own traces, so model and harness fit hand in glove.
The raw intelligence comes from a leading American-made frontier model: world-class capability, hosted and controlled in North America.
For high-stakes work, a second agent challenges and checks the result before it’s delivered.
Every Wasabe agent already runs inside Spice Forge. Claim your seat and hand it your worst Tuesday.