NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD84EG4VXK2JQJ36R79YR2V
Run Details
Model:
MoonshotAI: Kimi K3 (Reasoning, Low)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
37.1s
Cost:
$0.018479
Provider:
Moonshot AI
Evaluation summary
Score
Evaluator
Details Link
99.98%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
99.9920%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly crouched behind the bounce house, blueprint flapping against her knees. Bill wore a clown nose over his balaclava, which defeated the purpose and somehow made it worse. Across the lawn, thirty six-year-olds attacked a piñata shaped like a banker while parents sipped lemonade and pretended not to know each other. “The presents are under the big unicorn,” Molly whispered. Bill nodded solemnly. “And the unicorn is extremely judgmental today. Please hurry.” Their plan was elegant in the way wet cardboard is elegant: enter as magicians, distract the crowd with scarves, load gifts into a cake trolley, exit before the mascot arrived. Molly had brought fake mustaches, a priest collar, and three kazoos for morale. Bill had brought a real ferret named Kevin for reasons he refused to explain, though Kevin had his own tiny harness and looked more prepared than either of them. A balloon popped. Three toddlers screamed. Kevin escaped toward the snack table wearing Bill’s watch. “No heroics,” Molly said. “Grab, roll, vanish.” Molly sighed, checked her stopwatch, adjusted her mustache, and stepped into the sunlight with a criminal smile. The party had no idea generosity was about to be burgled by two idiots and one morally flexible ferret.
created_at
"2026-07-25T18:20:53.641Z"
7/25/2026, 6:20:53 PM
model
"moonshotai/kimi-k3-20260715"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
2065
moderation_latency
(null)
generation_time
37003
tokens_prompt
75
tokens_completion
1304
native_tokens_prompt
147
native_tokens_completion
1229
native_tokens_completion_images
(null)
native_tokens_reasoning
942
native_tokens_cached
147
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.0184791
router
(null)
provider_responses
0
endpoint_id
"650c9b12-7b6d-444d-8d92-df8daef84d7a"
id
"chatcmpl-6a64fe7429361060d39e5d0c"
is_byok
false
latency
2061
model_permaslug
"moonshotai/kimi-k3-20260715"
provider_name
"Moonshot AI"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785003653-r66AmwQkcKFOpiYmBHeT"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785003653-oMaLvvmIJB1TH9VnCFW7"
upstream_id
"chatcmpl-6a64fe7429361060d39e5d0c"
total_cost
0.0184791
cache_discount
0.0003969
upstream_inference_cost
0
provider_name
"Moonshot AI"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
99.98%
Matches word count
n/a
neededClean
false
words
202
100.00%
Dialogue to Total Word Ratio
Ratio: 10.29%, Deviation: 0.29%
neededClean
false
wordsTotal
204
wordsDialogue
21
99.9920%