Kimi K2.7-Code: Open-Source Coding Model with Better Token Efficiency

Moonshot AI has released Kimi K2.7-Code, an open-source coding model available on Hugging Face under the moonshotai/Kimi-K2.7-Code namespace. The model is tagged as image-text-to-text and uses the Transformers library. It positions itself as a token-efficient alternative for code generation and understanding tasks.
Key Features
- Inference providers: Novita offers the model with live status, tool calling support (
toolCalling: true), and structured output currently unavailable. Throughput measured at 36.1 tokens/second. - Model architecture: The model comes in 64 shards (safetensors format:
model-00001-of-000064.safetensors). - Token efficiency: The model uses a custom chat template that preserves reasoning content (
preserve_thinking: true) and optimizes token usage by separating history and suffix messages. The template includes special tokens like<|im_user|>,<|im_assistant|>, and<|im_system|>for role management, and<think>/</think>blocks to encapsulate chain-of-thought reasoning. - Tool calling: Native support for tool calls with structured argument formatting, using
<|tool_call_begin|>and<|tool_call_end|>markers. - Community engagement: 334 likes on Hugging Face, with 4 HN comments and 41 points as of publication.
Practical Implications
The template design explicitly avoids embedding reasoning tokens in history when preserve_thinking is false, reducing context overhead. For developers using AI coding agents, this means lower token consumption per interaction — especially beneficial for long agentic loops where reasoning chains are repeated. The tool calling format is JSON-aligned, making it straightforward to integrate with existing function-calling pipelines.
The model is available for immediate use via Novita, and the Hugging Face repository includes full tokenizer config and template source.
📖 Read the full source: HN AI Agents
👀 See Also

Anthropic study reveals cognitive degradation in AI-assisted workflows
Anthropic's global study of 80,000 users found academic users report cognitive degradation rates 2.5x higher than average when using AI tools like Claude and Cursor. The source identifies the problem as users eliminating the 'digestion phase' of work.

OpenClaw 2026.4.29 Broken – Downgrade to 2026.2.6
OpenClaw version 2026.4.29 is broken with random errors, slow CLI, double replies. Downgrade to 2026.2.6 to fix.

Spotify Rolls Out 'Verified' Badges to Tag Human Artists vs AI-Generated Acts
Spotify adds a green checkmark 'Verified by Spotify' badge to artist profiles that meet criteria like linked social accounts, concert dates, or merchandise, aiming to distinguish human acts from AI-generated ones.

DeepSeek Paid API Uses Prompts for Training — What OpenClaw Users Need to Know
DeepSeek's official API logs prompts for training, even on paid tiers. Gemini only logs on free AI Studio. OpenClaw now defaults to DeepSeek V4 Flash — beware when processing personal data.