NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36W55ZYSNEZ5BKPWCBNACB7
Run Details
Model:
Grok 4.7 (Reasoning, High)
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
1m 42s
Cost:
$0.041488
Provider:
xAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
99.9995%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly and Bill crouched in the shrubs beside the rental bouncy castle, watching parents arrange paper plates. A clown tested a horn. The present table sagged under glossy bags. "We move when the clown starts his trick," Molly whispered. "Everybody looks at him. Nobody looks at the gifts." Bill picked a leaf off his sleeve. He had insisted on the bow tie, which made him look like a guilty waiter. "This is a children's party, Molly. We are stealing toys from a kid named probably Tim, and I hate it." "Tim can get new toys," she said. "We cannot pay rent with feelings." They had walked the block twice. The side gate latch was loose. Caterer coats hung by the garage. "Aprons on, bags out, wagon waiting," she said. "If anyone stops us, we are fetching more ice and smiling very politely today, too." Bill watched a small girl hug a wrapped box. His mouth did its guilty thing. "We leave one present. The socks. A person has to sleep at night, Molly." Molly considered the yard, the cake, the easy pile. "Fine. Socks stay. Everything else comes with us, and we do not get sentimental about action figures."
created_at
"2026-09-23T10:14:37.316Z"
9/23/2026, 10:14:37 AM
model
"x-ai/grok-4.7-20260916"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
2467
moderation_latency
(null)
generation_time
101937
tokens_prompt
75
tokens_completion
746
native_tokens_prompt
1303
native_tokens_completion
8497
native_tokens_completion_images
(null)
native_tokens_reasoning
8243
native_tokens_cached
1152
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.041488
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790158477-3n8IQwFbMLtCC1VMg5Ph"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790158477-tO9mIyAQNettHInnLPIS"
upstream_id
"c11c923e-f3f2-9c0f-833d-6a416b481040"
provider_responses
0
endpoint_id
"c942f37a-87e1-4fa8-83d0-8c67fda4fb4a"
id
"c11c923e-f3f2-9c0f-833d-6a416b481040"
is_byok
false
latency
165
model_permaslug
"x-ai/grok-4.7-20260916"
provider_name
"xAI"
status
200
total_cost
0.041488
cache_discount
0.0013824
upstream_inference_cost
0
provider_name
"xAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
199
100.00%
Dialogue to Total Word Ratio
Ratio: 50.25%, Deviation: 0.25%
neededClean
false
wordsTotal
199
wordsDialogue
100
99.9995%