NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36WZ1X9EW4NW6GD8XTJ769R
Run Details
Model:
Grok 4.7 (Reasoning, High)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
1m 33s
Cost:
$0.035954
Provider:
xAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly spread the party invitation across the table as if it were a classified map. Balloons in the margin. A dinosaur piñata. A handwritten note about juice. Bill leaned over it, solemn, holding a spoon he swore was sentimental. "We walk in smiling," Molly said. "We leave with the presents." "Even the socks?" Bill asked. "Especially the socks. Leave the cake." The rest of the scheme was mostly confidence and bad lighting. They would arrive as backup entertainers, clap at the wrong moments, and orbit the gift table until orbit became ownership. Molly called this technique hospitality. Bill called it Tuesday. Neither mentioned tools, timers, or talent, because those would have ruined the joke and also required them to own any. They argued, briefly, about whether a piñata counted as security or as a coworker with commitment issues. Molly won by refusing to negotiate with papier-mache. The plan stayed gloriously unserious. On purpose. Outside, a neighbor's trumpet practiced the birthday song with criminal enthusiasm. Inside, Molly labeled a laundry basket DONATIONS in letters large enough to embarrass a dictionary. Bill stared at the basket as if it might confess first. "Ready?" he asked. "Never," Molly said, and stood up anyway.
created_at
"2026-09-23T10:28:45.104Z"
9/23/2026, 10:28:45 AM
model
"x-ai/grok-4.7-20260916"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
591
moderation_latency
(null)
generation_time
92357
tokens_prompt
75
tokens_completion
563
native_tokens_prompt
1303
native_tokens_completion
7344
native_tokens_completion_images
(null)
native_tokens_reasoning
7084
native_tokens_cached
1152
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0359536
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790159325-jEICfN2vIHIAWbKfcOOB"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790159325-cBmYT51gYmv2rXxxXaDf"
upstream_id
"07972b9d-3c67-953b-adcb-a61880a666b5"
provider_responses
0
endpoint_id
"48cfe052-27db-4511-bd21-db8053f20023"
id
"07972b9d-3c67-953b-adcb-a61880a666b5"
is_byok
false
latency
172
model_permaslug
"x-ai/grok-4.7-20260916"
provider_name
"xAI"
status
200
total_cost
0.0359536
cache_discount
0.0013824
upstream_inference_cost
0
provider_name
"xAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 9.95%, Deviation: 0.05%
neededClean
false
wordsTotal
201
wordsDialogue
20
100.0000%