NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD53VG2ZYHCZ7FENVA30QXW
Run Details
Model:
MoonshotAI: Kimi K3 (Reasoning, High)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
55.0s
Cost:
$0.028424
Provider:
Moonshot AI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly crouched behind the hedge, binoculars fogging as balloons bobbed beyond the picket fence. Bill wore a clown nose he insisted was tactical. Inside, toddlers hammered cake like tiny debt collectors. “We grab gifts, not glory,” Molly whispered. Bill nodded, then stepped on a kazoo. The yard erupted; parents turned; a bouncy castle sighed like an accusation. Molly’s blueprint fluttered into the punch bowl, where a pirate napkin sailed across it bravely. “Abort?” Bill squeaked. “Absolutely,” said Molly, already smiling. They retreated through sprinklers, clutching nothing but dignity, which had holes. Behind them, a girl offered Bill a sticker for being funny. He took it, solemn as a knight receiving treasure. By the van, Molly laughed so hard the maps shook. “New plan,” she said. Bill waited. “We return every present, become magicians, and get paid in cake.” The nose squeaked once, approving. Above them, the party blazed innocent and loud, while two villains practiced disappearing into decency, badly, gloriously, and just in time for the clown car of consequences to parallel park outside. Neighbors applauded, assuming performance art, while the birthday knight decreed stickers legal tender and Bill pocketed one like bail money from tomorrow, softly, no sirens yet.
created_at
"2026-07-25T17:28:08.457Z"
7/25/2026, 5:28:08 PM
model
"moonshotai/kimi-k3-20260715"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
2818
moderation_latency
(null)
generation_time
54932
tokens_prompt
75
tokens_completion
1841
native_tokens_prompt
147
native_tokens_completion
1892
native_tokens_completion_images
(null)
native_tokens_reasoning
1591
native_tokens_cached
147
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.0284241
router
(null)
provider_responses
0
endpoint_id
"650c9b12-7b6d-444d-8d92-df8daef84d7a"
id
"chatcmpl-6a64f2292f0a87cfa2ae5023"
is_byok
false
latency
2817
model_permaslug
"moonshotai/kimi-k3-20260715"
provider_name
"Moonshot AI"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785000488-2awRo5IYCLNAkPrNPvRa"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785000488-LugHb2uQovZSc5hfOnyZ"
upstream_id
"chatcmpl-6a64f2292f0a87cfa2ae5023"
total_cost
0.0284241
cache_discount
0.0003969
upstream_inference_cost
0
provider_name
"Moonshot AI"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 10.00%, Deviation: 0.00%
neededClean
false
wordsTotal
200
wordsDialogue
20
100.0000%