NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD6BTW6H3JNXQCD6NJHH3T2
Run Details
Model:
MoonshotAI: Kimi K3 (Reasoning, High)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
1m 55s
Cost:
$0.062414
Provider:
Moonshot AI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly and Bill crouched behind a hedge shaped like a giraffe, studying the Henderson backyard through opera glasses that kept fogging from dread and pudding. Bunting giggled in the breeze. Somewhere inside, a birthday magician applauded himself for surviving another rabbit. Molly whispered, "We slip in as clowns, grab gifts, leave before cake." Bill said, "Brilliant, except I look honest in sequins and moral doubt." Molly unfolded a map drawn in ketchup, which Bill mistook for evidence and apologized to. Their disguises were two clown suits, one tutu, and a banner reading Congratulations Graduate Kevin. No one at a six year old party would question Kevin, unless Kevin arrived. They practiced innocent faces in a spoon, then blamed the spoon for looking guilty. A piƱata donkey watched them with papery menace, as if it knew their credit scores. Bill suggested tunneling, noticed the lawn was sprinkler, and chose interpretive sneaking instead. Molly checked the schedule: cake at two, chaos at two oh one, remorse by two oh five. When the doorbell rang, every child shouted strangers, the dog opened its gift, and Grandma Pat arrived carrying nunchucks labeled book club, so our heroes reconsidered crime and joined the conga line.
created_at
"2026-07-25T17:49:58.54Z"
7/25/2026, 5:49:58 PM
model
"moonshotai/kimi-k3-20260715"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
2252
moderation_latency
(null)
generation_time
114734
tokens_prompt
75
tokens_completion
4048
native_tokens_prompt
147
native_tokens_completion
4158
native_tokens_completion_images
(null)
native_tokens_reasoning
3875
native_tokens_cached
147
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.0624141
router
(null)
provider_responses
0
endpoint_id
"650c9b12-7b6d-444d-8d92-df8daef84d7a"
id
"chatcmpl-6a64f747afdc8cdbbb9b5e8d"
is_byok
false
latency
2252
model_permaslug
"moonshotai/kimi-k3-20260715"
provider_name
"Moonshot AI"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785001798-MFO3PON0vIxvK4moqron"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785001798-Nrxxcv7Uq5oD0s1hSgbH"
upstream_id
"chatcmpl-6a64f747afdc8cdbbb9b5e8d"
total_cost
0.0624141
cache_discount
0.0003969
upstream_inference_cost
0
provider_name
"Moonshot AI"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 10.00%, Deviation: 0.00%
neededClean
false
wordsTotal
200
wordsDialogue
20
100.0000%