Hamartia Antidote
Elite Member
Pentagon Expands AI Arsenal With Grok and ChatGPT for Military Use
The Pentagon is expanding its military AI platform with Grok and ChatGPT while keeping multiple models in the mix.
The Pentagon continued its artificial intelligence push this week, announcing on Monday that it adopted Starshield AI’s Grok for Government and OpenAI’s ChatGPT for unclassified military use. This further expands the number of AI assistants now in use with the Department of War (DoW).
Both Grok for Government and ChatGPT are accredited for Controlled Unclassified Information (CUI) at Impact Level 5 (IL5) and are engineered for secure, consistent enterprise use.
“Starshield AI’s Grok for Government will provide the warfighter with immediate productivity gains, stronger knowledge continuity and more secure and efficient collaboration. By bringing Starshield AI’s Grok to GenAI.mil, Department personnel will gain access to advanced capabilities, including deep-thinking inference, adaptive reasoning modes (Auto, Fast, and Expert), customizable workspaces, persistent projects, and reusable “playbooks” that capture and scale institutional knowledge,” the Pentagon explained.
It added that “ChatGPT Mil brings a familiar commercial experience into the Department’s secure environment, tailored to warfighter needs. The core experience centers on chat, files, projects, and custom GPTs, with additional features sequenced over time. ChatGPT Mil supports document-heavy unclassified work across the Department, including planning, policy, logistics, and administration, and is built for scale to support more than 3 million Department personnel.”
Both AI models will be offered via the Pentagon’s GenAI.mil platform, which launched in December 2025. It initially included only Google’s Gemini for Government.
“The Pentagon is doing the right thing by using different models and providers. We do not know who will be the provider of the best models, which makes it prudent to work with all of them,” said technology industry analyst Roger Entner, founder of Recon Analytics.
“It also allows them to have a ‘Council of Models’ where the different models from different providers make joint decisions, which is better than relying on just one model,” Entner told ClearanceJobs.
Move Away From Specialized Programs
The addition of Grok for Government and ChatGPT Mil is a move away from earlier Pentagon AI initiatives, which frequently remained concentrated in specialized programs, individual contracts, experimental units, innovation offices, and narrowly defined applications.That left considerable distance between an impressive demonstration and technology that changes how people throughout the department actually work, geopolitical analyst Irina Tsukerman, president of threat assessment firm Scarab Rising, told ClearanceJobs.
“Routine access to frontier models can close some of that distance by allowing personnel to discover applications through everyday use in intelligence research, logistics, procurement, acquisition, software development, planning, document exploitation, and administrative work,” Tsukerman added.
Grok for Government and ChatGPT Mil will also provide the Pentagon with a much larger body of evidence about where generative AI genuinely improves performance. Tsukerman explained that it will accelerate an existing process, which applications personnel actually find useful, yet errors could become harder to detect precisely because the resulting product looks polished.
“Using several major models can give DoD an opportunity to study their respective strengths, weaknesses, biases, and failure patterns under military conditions while reducing the danger of building an entire institutional AI environment around the assumptions and technical choices of one company,” she continued.
Moreover, ChatGPT and Grok can interpret identical information differently because their training, tuning, safeguards, architectures, and approaches to generating answers vary, and those divergences can expose assumptions or uncertainties that deserve human examination. In turn, the Pentagon could eventually develop systems that direct different types of tasks toward models demonstrated to perform particularly well in those areas. At the same time, analysts could deliberately compare outputs when dealing with ambiguous evidence or competing hypotheses.
“A military and intelligence establishment already accustomed to alternative analysis, red teams, and competing assessments has an institutional foundation for turning model disagreement into an analytical tool,” Tsukerman acknowledged.
“The Pentagon’s most useful measure of progress will therefore come from what happens to the quality of human judgment as machine capability expands,” she continued. “A military that learns how to question extremely capable AI systems can gain a substantial advantage from them; a military that gradually allows those systems to determine how problems are framed can become faster and more technologically sophisticated while making its own decision-making increasingly vulnerable to manipulation.”
