Can Kimi K3 disable reasoning?
K3 keeps reasoning continuously enabled. It is recommended to set reasoning_effort: max; if omitted, it is used this way as well. Do not copy the thinking switch from K2.6 into K3 requests. If the task is only brief rewriting or classification, consider a model better suited to lightweight tasks rather than relying on disabling reasoning.
Can K3 view screenshots and output webpage code?
You can submit screenshots as image_url image blocks together with framework, style specifications, and interaction requirements for interface analysis and code generation. It outputs text code and suggestions; it does not automatically run the webpage as a result. Verification should be completed through actual rendering, screenshot comparison, and interaction testing.
What content needs to be saved for multi-turn tool calls?
When using a dedicated endpoint, you should save and return the complete assistant message, including the returned reasoning_content, tool_calls, as well as the corresponding call IDs and tool results. Keeping only the final answer will lose execution state. When using a hosted session, continue the task through the same session ID.
How can K3 return results that programs can process?
You can explicitly request JSON output in the prompt and specify field names, types, required fields, and allowed values, with examples; the application should perform JSON parsing, structural validation, and business validation, and set up retries and exception handling for missing fields, type errors, or truncated results. If using response_format to configure the format, first verify the actual effect of the selected mode in kimi-k3 requests, and do not treat the json_schema field as a guarantee of strict Schema compliance.
How should materials be submitted for PDF analysis?
When a file-reading workflow is needed, you can use the AI Chat v2 file_url content block to provide an accessible file link; alternatively, extract the text first and then submit it to a dedicated endpoint for analysis. File reading is an endpoint feature and does not mean the model directly accepts arbitrary formats; the clarity of scanned content also affects analysis.