Back to Newsroom

Pokee AI Releases Pokee-Isaac 28B: A 10M-Token Context Agentic Model Built to Run Inside the Customer Boundary

By Modelverse Editorial·August 8, 2026·2 min read
Pokee AI Releases Pokee-Isaac 28B: A 10M-Token Context Agentic Model Built to Run Inside the Customer Boundary

Pokee AI has unveiled Pokee-Isaac 28B, a significant new 28-billion parameter, text-only foundation model featuring an unprecedented 10-million token context window. What truly sets Isaac apart is its design philosophy: it's built to operate entirely within the customer's security perimeter, whether deployed in a Virtual Private Cloud (VPC), on-premises, or directly on-device. This addresses a critical challenge for long-horizon AI agents, which rapidly accumulate context but have historically relied on cloud endpoints, making them unsuitable for regulated industries, public sector institutions, and applications with strict data residency requirements.

This innovative approach allows organizations to leverage advanced AI agent capabilities without compromising data privacy or security. Pokee-Isaac is offered through a licensed, OpenAI-compatible developer API, enabling seamless integration into existing workflows. By keeping data within the customer boundary, the model empowers sensitive applications to harness the power of extensive context understanding and coherent reasoning, a capability previously almost exclusive to external cloud services.

For developers and researchers, Pokee-Isaac 28B represents a breakthrough in deploying powerful, long-context AI agents in data-sensitive environments. The model demonstrates robust performance, achieving 93.3% on RULER at its full 10M-token context and showing parity with strong cloud baselines on agentic benchmarks like BFCL v4 and τ³-bench. Crucially, it's engineered for efficiency, capable of running on a single GPU (such as an RTX 4090 or B200-class), with Day-0 support for vLLM and SGLang, making high-performance, in-boundary AI agents a practical reality.

ai-newsbreakingmarktechpost

Footnotes & Primary References

Related content

Writer introduces new AI model and upgraded harness to contain token costs

Built as a post-training variation on Z.ai's open source model GLM-5.2, Writer says the new system should provide deployment-ready capabilities at a much lower price.

Read article

OpenAI introduces 'Ultrafast,' a new mode that makes GPT-5.6 Sol work at 14x the speed

OpenAI is launching a preview of a sped up version of its latest, most powerful model, in an effort to court enterprise users.

Read article

Nvidia's new $500B plan is risky but brilliant, especially for aging GPUs

Nvidia has a plan to make sure its GPUs won't lose value. It wants to convince a new crop of financiers to keep lending for AI buildouts.

Read article