NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD5Y66BVPPMQ4EFRPQR423F
Run Details
Model:
MoonshotAI: Kimi K3 (Reasoning, Low)
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
13.4s
Cost:
$0.005114
Provider:
Moonshot AI
Evaluation summary
Score
Evaluator
Details Link
95.99%
Matches word count
(details)
0.00%
Dialogue to Total Word Ratio
(details)
47.9934%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly spread the blueprints across the hood of the getaway minivan. "The Henderson birthday party. Saturday, two o'clock. Forty presents, minimum." "Forty?" Bill squinted at the diagram. "What's the security situation?" "One clown, possibly two. A bouncy castle with structural integrity issues. And a piƱata we can use as a distraction." "I don't like clowns, Molly." "Nobody likes clowns, Bill. That's why they're perfect cover. You go in dressed as one, nobody makes eye contact with you." "What about the cake? Do we take the cake?" "We are not animals. We leave the cake." "But it's chocolate." "Bill. Focus. The presents get loaded into the bouncy castle, we deflate it, roll it out like a burrito, and we're gone before musical chairs ends." "And the kids? Won't they scream?" "That's what the puppet show is for. You do the puppet show." "I thought I was the clown!" "You're a clown who does puppet shows. Range, Bill. That's what separates us from amateurs." Bill sighed, pulling a red nose from his pocket. "One of these days we'll rob a bank like normal criminals." "Banks don't have goodie bags, Bill. Think of the margins."
created_at
"2026-07-25T17:42:31.388Z"
7/25/2026, 5:42:31 PM
model
"moonshotai/kimi-k3-20260715"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
2394
moderation_latency
(null)
generation_time
13346
tokens_prompt
75
tokens_completion
352
native_tokens_prompt
147
native_tokens_completion
338
native_tokens_completion_images
(null)
native_tokens_reasoning
52
native_tokens_cached
147
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.0051141
router
(null)
provider_responses
0
endpoint_id
"650c9b12-7b6d-444d-8d92-df8daef84d7a"
id
"chatcmpl-6a64f588689654d454ea00d0"
is_byok
false
latency
2394
model_permaslug
"moonshotai/kimi-k3-20260715"
provider_name
"Moonshot AI"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785001351-ZqSj8WMNCc45bBe2cIXI"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785001351-3hRqQCK2hrT3or3dYpQ0"
upstream_id
"chatcmpl-6a64f588689654d454ea00d0"
total_cost
0.0051141
cache_discount
0.0003969
upstream_inference_cost
0
provider_name
"Moonshot AI"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
95.99%
Matches word count
n/a
neededClean
false
words
192
0.00%
Dialogue to Total Word Ratio
Ratio: 86.98%, Deviation: 36.98%
neededClean
false
wordsTotal
192
wordsDialogue
167
47.9934%