NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36SPPMP4GX09QDDX62M28XM
Run Details
Model:
Grok 4.7 (Reasoning, High)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
1m 7s
Cost:
$0.028235
Provider:
xAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly peered through a hedge that smelled like wet dog and cancelled plans. Beyond it, the Henderson birthday sprawled across the lawn, an inflatable castle slumping as if it already regretted the invitation. Bill had parked the van half on the curb and half on his conscience. On the dash sat a grocery receipt he had labeled TARGETS, the letters so proud they nearly punched through the paper. Cupcakes. A red scooter. Three boxes in dinosaur wrap. They had brought no gloves, no map, and no excuse a reasonable adult would accept. Squeals drifted over the fence, chased by one heroic kazoo. A clown dropped the cake. Molly decided the universe was already cooperating. Bill tapped the wheel. They had agreed, in the solemn way of a meeting held in a ditch, that presents left on a picnic table were unattended and therefore negotiable. "Front door or the castle?" Molly asked. "The castle," Bill said. "Fewer witnesses with dignity." "If a parent sees us?" "We are the entertainment." She nodded, which passed for their whole strategy, and lifted the kazoo she had stolen from a yard sale back in 2009, already ready to announce a small crime nobody had scheduled.
created_at
"2026-09-23T09:31:45.692Z"
9/23/2026, 9:31:45 AM
model
"x-ai/grok-4.7-20260916"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
555
moderation_latency
(null)
generation_time
67045
tokens_prompt
75
tokens_completion
522
native_tokens_prompt
1303
native_tokens_completion
5736
native_tokens_completion_images
(null)
native_tokens_reasoning
5483
native_tokens_cached
1152
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0282352
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790155905-6d63pfXve8r4BODOevbs"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790155905-fLjI2aJqZfEfMaK8AgWh"
upstream_id
"b19225bd-5489-940f-9bd6-bad77bc784bf"
provider_responses
0
endpoint_id
"c942f37a-87e1-4fa8-83d0-8c67fda4fb4a"
id
"b19225bd-5489-940f-9bd6-bad77bc784bf"
is_byok
false
latency
181
model_permaslug
"x-ai/grok-4.7-20260916"
provider_name
"xAI"
status
200
total_cost
0.0282352
cache_discount
0.0013824
upstream_inference_cost
0
provider_name
"xAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 10.00%, Deviation: 0.00%
neededClean
false
wordsTotal
200
wordsDialogue
20
100.0000%