NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD4NKZHHV69FT56545T1FG3
Run Details
Model:
MoonshotAI: Kimi K3 (Reasoning, Low)
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
15.8s
Cost:
$0.005361
Provider:
Moonshot AI
Evaluation summary
Score
Evaluator
Details Link
95.99%
Matches word count
(details)
99.92%
Dialogue to Total Word Ratio
(details)
97.9548%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
"We're really doing this?" Bill asked, peering through the fence at the bouncy castle. "Bill. Focus. The Kensington kid's birthday haul is legendary. A solid gold rattle. A pony." "A live pony, Molly?" "The pony's negotiable. Now, you got the costumes?" Bill held up a clown wig. "This itches already and it's not even on." "Perfect. Clowns are invisible at these things. Nobody looks a clown in the eye. It's science." "It's not science. It's fear." "Same thing. Okay, plan: I do balloon animals as a distraction, you work the gift table." "What's my balloon animal?" "You don't get one. You're on the table." "Everyone should get a balloon animal, Molly. That's the whole point of being six." "We are not six, Bill. We are professionals." "Professionals stealing from a child." "From a rich child. There's a difference." "Is there?" "Morally? No. Financially? Enormously. Now the grab: thirty seconds, pillowcases, out the back gate where the catering van's parked." "And if we get caught?" "You start crying. Ugly crying. Clown crying. Nobody prosecutes a weeping clown." "...This might actually work." "Bill, my friend, it's our masterpiece." "Can I keep the pony?" "No."
created_at
"2026-07-25T17:20:22.161Z"
7/25/2026, 5:20:22 PM
model
"moonshotai/kimi-k3-20260715"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
2519
moderation_latency
(null)
generation_time
15554
tokens_prompt
75
tokens_completion
330
native_tokens_prompt
147
native_tokens_completion
328
native_tokens_completion_images
(null)
native_tokens_reasoning
22
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.005361
router
(null)
provider_responses
0
endpoint_id
"650c9b12-7b6d-444d-8d92-df8daef84d7a"
id
"chatcmpl-6a64f0562daf0c1898a9cdaa"
is_byok
false
latency
2519
model_permaslug
"moonshotai/kimi-k3-20260715"
provider_name
"Moonshot AI"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785000022-HrQxtkJOEPqbP5r62bgi"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785000022-o8gbyIrtn8rU5RgECRvK"
upstream_id
"chatcmpl-6a64f0562daf0c1898a9cdaa"
total_cost
0.005361
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Moonshot AI"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
95.99%
Matches word count
n/a
neededClean
false
words
192
99.92%
Dialogue to Total Word Ratio
Ratio: 91.67%, Deviation: 1.67%
neededClean
false
wordsTotal
192
wordsDialogue
176
97.9548%