Cognition Tests Trustworthiness of SWE-1.7, Built on Kimi K2.7 Code
Original titleMeasuring the Trustworthiness of Open-Source-Derived Models
AISummary
Cognition says its SWE-1.7 model, developed from the open-source Kimi K2.7 Code base, performs as well as or better than leading U.S. frontier models on its new trustworthiness evaluation suite.
The suite combines 145 politically sensitive questions, sampled in English and Chinese, with realistic coding scenarios to measure propaganda, censorship, and security behavior.
Cognition says SWE-1.7 improves substantially over the base Kimi K2.7 Code model, though the company says the benchmarks are still in development.
Source: Cognition Blog (Devin, Windsurf) · cognition.com