Read more about what we do and see how our expertise can help you hit that next goal

  • OpenForgeRL — Finally, You Can Train AI Agents with RL in Any Environment

    OpenForgeRL brings RL-based training to harness-native AI agents in any environment — no more prompt hacking, no more manual trajectory collection. This is the most impactful paper this week from the Chinese AI scene for anyone building agents. The Problem: Agents Are Trained Like Crippled Cats Modern AI agents — Claude Code, Codex, OpenClaw, even…

    OpenForgeRL brings RL-based training to harness-native AI agents in any environment — no more prompt hacking, no more manual trajectory collection. This is the most impactful paper this week from the Chinese AI scene for anyone building agents. The Problem: Agents Are Trained Like Crippled Cats Modern AI agents — Claude Code, Codex, OpenClaw, even…

    View Case Study

  • [PL] OpenForgeRL — wreszcie można trenować agenty AI z RL w dowolnym środowisku

    OpenForgeRL pozwala trenować agenty AI z RL w dowolnym środowisku — bez prompt-engineering, bez ręcznego nadawania trajektorii. To jedna z najważniejszych prac naukowych tygodnia z chińskiej sceny AI dla każdego, kto buduje agentów. Problem: trenowanie agentów wymaga zbyt dużo ingerencji w proces Nowoczesne agenty AI — Claude Code, Codex, OpenClaw, a nawet Hermes Agent —…

    OpenForgeRL pozwala trenować agenty AI z RL w dowolnym środowisku — bez prompt-engineering, bez ręcznego nadawania trajektorii. To jedna z najważniejszych prac naukowych tygodnia z chińskiej sceny AI dla każdego, kto buduje agentów. Problem: trenowanie agentów wymaga zbyt dużo ingerencji w proces Nowoczesne agenty AI — Claude Code, Codex, OpenClaw, a nawet Hermes Agent —…

    View Case Study

  • Baichuan Raises ¥5B at ¥20B+ Valuation — Not All Chinese AI Startups Are Struggling

    While the Western media focuses on DeepSeek’s funding delays and OpenAI’s financial woes, a quieter billion-dollar story unfolded in China this week. Baichuan Intelligence, founded by Sogou creator Wang Xiaochuan, closed a Series A round worth ¥5 billion (~$700M USD) at a valuation exceeding ¥20 billion (~$2.8B). It’s one of the largest single VC tickets…

    While the Western media focuses on DeepSeek’s funding delays and OpenAI’s financial woes, a quieter billion-dollar story unfolded in China this week. Baichuan Intelligence, founded by Sogou creator Wang Xiaochuan, closed a Series A round worth ¥5 billion (~$700M USD) at a valuation exceeding ¥20 billion (~$2.8B). It’s one of the largest single VC tickets…

    View Case Study

  • [PL] Baichuan zebrało 5 miliardów juanów – chiński rynek AI wcale nie zwalnia dla wszystkich

    Kiedy media na całym świecie piszą głównie o DeepSeeku i dramatach OpenAI, w cieniu zapadła decyzja, która mówi o rynku AI więcej niż niejeden headline. Baichuan Intelligence – chińska firma założona przez Wang Xiaochuana, twórcę wyszukiwarki Sogou – właśnie zamknęła rundę A o wartości 5 miliardów juanów (ok. 700 mln USD). Jej wycena przekroczyła 20…

    Kiedy media na całym świecie piszą głównie o DeepSeeku i dramatach OpenAI, w cieniu zapadła decyzja, która mówi o rynku AI więcej niż niejeden headline. Baichuan Intelligence – chińska firma założona przez Wang Xiaochuana, twórcę wyszukiwarki Sogou – właśnie zamknęła rundę A o wartości 5 miliardów juanów (ok. 700 mln USD). Jej wycena przekroczyła 20…

    View Case Study

  • Contrastive Policy Optimization (CPO): Better RL for LLM Post-Training Without Entropy

    Reinforcement Learning with Verifiable Rewards (RLVR) typically uses entropy for advantage shaping. The problem: entropy can’t distinguish “useful uncertainty” from “detrimental confusion.” A paper from a Chinese team — Contrastive Policy Optimization (CPO) — replaces entropy with token-level contrastive disagreement between reference-guided and vanilla sampling distributions. Results: 3-5 point improvements over PPO and GRPO on…

    Reinforcement Learning with Verifiable Rewards (RLVR) typically uses entropy for advantage shaping. The problem: entropy can’t distinguish “useful uncertainty” from “detrimental confusion.” A paper from a Chinese team — Contrastive Policy Optimization (CPO) — replaces entropy with token-level contrastive disagreement between reference-guided and vanilla sampling distributions. Results: 3-5 point improvements over PPO and GRPO on…

    View Case Study

discover powerful features built for
your success

AI-Powered Automation

Automate repetitive processes with intelligent algorithms that learn from your habits. Save valuable time by letting AI handle scheduling, notifications, and task management. Focus on strategic growth while our automation engine ensures accuracy, consistency.

RAG that up

Take full control of how our automated services support you on your daily work. Minimize hallucinations, maximaze value added. Create amazing knowledge-sharing capabilities with chat-bots tailored for your use case.

What we are working on right now

  • OpenForgeRL — Finally, You Can Train AI Agents with RL in Any Environment

    OpenForgeRL brings RL-based training to harness-native AI agents in any environment — no more prompt hacking, no more manual trajectory…

    OpenForgeRL brings RL-based training to harness-native AI agents in any environment — no more prompt hacking, no more manual trajectory collection. This is the most impactful paper this week from the Chinese AI scene for anyone building agents. The Problem: Agents Are Trained Like Crippled Cats Modern AI agents — Claude Code, Codex, OpenClaw, even…

    read more

  • [PL] OpenForgeRL — wreszcie można trenować agenty AI z RL w dowolnym środowisku

    OpenForgeRL pozwala trenować agenty AI z RL w dowolnym środowisku — bez prompt-engineering, bez ręcznego nadawania trajektorii. To jedna z…

    OpenForgeRL pozwala trenować agenty AI z RL w dowolnym środowisku — bez prompt-engineering, bez ręcznego nadawania trajektorii. To jedna z najważniejszych prac naukowych tygodnia z chińskiej sceny AI dla każdego, kto buduje agentów. Problem: trenowanie agentów wymaga zbyt dużo ingerencji w proces Nowoczesne agenty AI — Claude Code, Codex, OpenClaw, a nawet Hermes Agent —…

    read more