NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36T4HPBQ7V99J34D7FW633E
Run Details
Model:
Grok 4.7 (Reasoning, High)
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
1m 24s
Cost:
$0.035430
Provider:
xAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly crouched behind the plastic hedge and watched parents drift toward the snack table like sleepy cattle. The bounce house wheezed. A clown practiced a sad balloon dog by the gate. Streamers tugged at the fence in a warm, useless wind. Beside her, Bill kept adjusting a mustache that had already given up. "Bill, we go in when the piƱata breaks," Molly said. "Everyone stares at candy. Nobody stares at gifts." Bill nodded at the folding table stacked with bright paper. "Wrapped boxes, no locks, parents drunk on juice. This might be our easiest afternoon." She counted the exits: side gate, kitchen door, and a gap behind the inflatable castle. "In, grab the presents, out before the clown finishes that dog. If anyone asks, we are backup entertainment." "What if that clown asks us first?" "Then we clap. Clowns love clapping." Molly checked her watch and lowered her voice. "Party starts in ten. Try to look like someone who enjoys children." Bill considered the unattended pile and failed to look innocent. "I enjoy their gifts more than I enjoy children." "Close enough, Bill," she said. "Do not eat the cake this time. Leave the balloons. We are thieves, not monsters."
created_at
"2026-09-23T09:39:19.377Z"
9/23/2026, 9:39:19 AM
model
"x-ai/grok-4.7-20260916"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
599
moderation_latency
(null)
generation_time
83996
tokens_prompt
75
tokens_completion
906
native_tokens_prompt
1303
native_tokens_completion
7267
native_tokens_completion_images
(null)
native_tokens_reasoning
7000
native_tokens_cached
1280
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0354304
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790156359-yUJ8YwalHQx4ThLtx4nd"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790156359-1UJBnxI7G3sZ1EsztlIy"
upstream_id
"c8bac706-e151-9615-b233-4bc615c3f959"
provider_responses
0
endpoint_id
"c942f37a-87e1-4fa8-83d0-8c67fda4fb4a"
id
"c8bac706-e151-9615-b233-4bc615c3f959"
is_byok
false
latency
171
model_permaslug
"x-ai/grok-4.7-20260916"
provider_name
"xAI"
status
200
total_cost
0.0354304
cache_discount
0.001536
upstream_inference_cost
0
provider_name
"xAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 50.00%, Deviation: 0.00%
neededClean
false
wordsTotal
200
wordsDialogue
100
100.0000%