Verification
Independent proof that a result is actually correct.
sagtuku. An AI + X neo lab by Alpai GmbH.
We build AI that takes on real goals, does the work, checks the results, and learns from what actually happened. Not from what it expected.
01The missing piece
Today’s best AI can reason, write code, and use tools. But to work on its own, it still lacks five basic skills.
“The next step is teaching machines consequences.”
Tell whether an action really worked.
Stay on one goal for days, weeks, or months.
Get better from real experience.
Foresee what an action will cause.
Reuse skills in new settings, from code to machines.
02The closed loop
Most AI is trained once and then stays the same. Our systems run in a loop. Every round adds experience and evidence to the next.
Understand the goal, the limits, and what success looks like.
Do the work with approved tools and code.
Look at what really changed. Don't assume the plan worked.
Check the result with tests the agent does not control.
Record what happened, including failures and sources.
Turn checked experience into better skills.
Start the next task knowing more than the last.
03What we research
Models are getting smarter. We work on what they still need to finish real tasks: staying on track, checking results, and learning from mistakes.
Independent proof that a result is actually correct.
Agents that stay on track through long, complex tasks.
Turn successes and failures into lessons for future work.
Predict what actions will cause, and carry skills into new settings.
04Phase I · Research agent
Our first product is an AI agent for technical research and engineering. We start with digital work because every step can be inspected and repeated.
Find out if architecture X can cut inference costs without hurting performance. Design and run the experiments, analyze the results, question your own conclusions, and show your work so others can repeat it.
What you get back
05How we measure success
How long can the AI work on a real goal before a person has to step in?
How many finished tasks pass an independent check?
06Questions
sagtuku is an AI + X neo lab operated by Alpai GmbH, a Swiss AI company. "AI + X" means we pair AI with a real field of work (X), such as software, science, or engineering, and build systems that deliver results in that field.
It is AI that does not stop at an answer. It acts, looks at what happened, checks the result, and remembers the lesson. Each task makes it better at the next one.
An AI agent for technical work: software engineering, AI research, data analysis, experiment design, and model evaluation. We start here because every step and result in these fields can be checked and repeated.
It does not take its own word for it. Separate checks test the result: automated tests, simulations, critical reviews, statistics, or repeated experiments. When the evidence is unclear, a person makes the call.
No. We build the system around the model: long-running tasks, independent checks, memory, and learning from experience. The goal is AI you can trust to finish real work, not a new chat window.
↗Work with us
Working on a hard research or engineering question? Start with the result you need to prove.