NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36SEK0JHF05ADQ8TB00G9F9
Run Details
Model:
Grok 4.7 (Reasoning, High)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
1m 8s
Cost:
$0.026344
Provider:
xAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly crouched behind the hedge like a shrub with ambition, binoculars fixed on a bounce house that sighed under six screaming children. Bill adjusted a fake mustache that had already quit. Beyond the fence, a pony glared beside a tower of presents wrapped in paper bright enough to signal ships. Their whole plan depended on cake and terrible timing. "We move only when they cut the cake," Molly whispered. Bill swallowed hard. "I am not wrestling that clown." She tapped a napkin map stained with lunch: side gate, past the piƱata, snatch the pile, vanish before anyone followed the ribbon trail. "And do not pet the pony," she added. He nodded, which covered most of his duties. A balloon popped and both of them flinched like experts. They compared watches that disagreed by eleven minutes, then agreed time was optional. The presents waited in a shiny heap, innocent and already doomed. Somewhere a kazoo launched the worst song either thief had ever personally financed. Molly lowered the binoculars. The pony had not blinked. Bill shifted the mustache to the other cheek, hoping disguise was mostly confidence. Neither of them mentioned that the gate was probably locked, since optimism beat equipment.
created_at
"2026-09-23T09:27:19.834Z"
9/23/2026, 9:27:19 AM
model
"x-ai/grok-4.7-20260916"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
778
moderation_latency
(null)
generation_time
67628
tokens_prompt
75
tokens_completion
563
native_tokens_prompt
1303
native_tokens_completion
5374
native_tokens_completion_images
(null)
native_tokens_reasoning
5130
native_tokens_cached
1280
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.026344
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790155639-I9IP0lLQeKl42aKnmrKD"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790155639-BhMK5ed25aEuR3VeBYMa"
upstream_id
"7125f5c0-7a6f-96ce-a854-36f3fe65a29a"
provider_responses
0
endpoint_id
"c942f37a-87e1-4fa8-83d0-8c67fda4fb4a"
id
"7125f5c0-7a6f-96ce-a854-36f3fe65a29a"
is_byok
false
latency
180
model_permaslug
"x-ai/grok-4.7-20260916"
provider_name
"xAI"
status
200
total_cost
0.026344
cache_discount
0.001536
upstream_inference_cost
0
provider_name
"xAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 10.00%, Deviation: 0.00%
neededClean
false
wordsTotal
200
wordsDialogue
20
100.0000%