NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36VZ9ZQ9DAD06BA15JS8KVV
Run Details
Model:
Grok 4.7 (Reasoning, High)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
1m 12s
Cost:
$0.030765
Provider:
xAI
Evaluation summary
Score
Evaluator
Details Link
99.74%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
99.8722%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly crouched behind the hedge with a clipboard, as if shrubbery required management. Bill adjusted a paper crown from the recycling and squinted toward the gift table like a man solving weather. Balloons knocked together. A plastic castle leaned. Wrapped boxes sat in a bright, unguarded heap while a toddler argued with a piƱata. The afternoon smelled like sugar and regret. Their plan was less a plan than a shared delusion. Molly wanted the cake song, that brief minute when every adult faced one direction and forgot the concept of watching things. Bill had contributed a laundry basket and a theory that confidence could replace competence. A clown honked past. Bill met the mulch. Molly did not blink. She had already noticed the mother orbiting the presents, and the dog, who had picked a side. Nobody watched them yet. Nearly time. "Wait for the candles," she said. "Then the basket," he said. "We grab the boxes and run." "Past the dog?" "Especially past the dog." The cake tilted into view. Small voices lifted. Molly counted ribbons. Bill gripped plastic handles and tried to become smaller than his own bad idea. A balloon escaped overhead. Behind them the bouncy castle breathed, patient as a verdict.
created_at
"2026-09-23T10:11:24.797Z"
9/23/2026, 10:11:24 AM
model
"x-ai/grok-4.7-20260916"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
531
moderation_latency
(null)
generation_time
72377
tokens_prompt
75
tokens_completion
606
native_tokens_prompt
1303
native_tokens_completion
6263
native_tokens_completion_images
(null)
native_tokens_reasoning
6005
native_tokens_cached
1152
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0307648
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790158284-nzgEQtDLACMJGNFV6Dak"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790158284-9q5dDDn14NANeCgyntYt"
upstream_id
"4dac7d59-577a-9d04-8baf-0d7a31dba04b"
provider_responses
0
endpoint_id
"c942f37a-87e1-4fa8-83d0-8c67fda4fb4a"
id
"4dac7d59-577a-9d04-8baf-0d7a31dba04b"
is_byok
false
latency
184
model_permaslug
"x-ai/grok-4.7-20260916"
provider_name
"xAI"
status
200
total_cost
0.0307648
cache_discount
0.0013824
upstream_inference_cost
0
provider_name
"xAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
99.74%
Matches word count
n/a
neededClean
false
words
204
100.00%
Dialogue to Total Word Ratio
Ratio: 9.80%, Deviation: 0.20%
neededClean
false
wordsTotal
204
wordsDialogue
20
99.8722%