NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36T2YYZHMSNJSNWF334M7M3
Run Details
Model:
Grok 4.7 (Reasoning, High)
Test:
Dialogue tags
Scenario:
Write unattributed dialogue
Duration:
7.4s
Cost:
$0.002413
Provider:
xAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Count dialogue tags
(details)
Model input
User/Human
You are an expert novel writer in the cozy mystery genre. Write a scene of about 200 words, with dialogue between two characters. The dialogue should be unattributed, meaning the speaker is not explicitly mentioned. The scene should be set in a small town bakery. Only return the story, nothing else.
Model response
Messages
Meta
AI
The bell over the door had barely settled when the smell of cinnamon reached the counter, warm and a little too sweet for the news being passed across it. "She never misses Thursday. Sourdough, no seeds, exact change in the blue coin purse." "I boxed it anyway. It's still in the back, getting hard around the edges." Flour dusted the glass case. Outside, rain tapped the awning, and Main Street looked rinsed and empty. "Her kitchen light was on all night. I saw it from the alley when I took the bins out." "Maybe she couldn't sleep." "Maybe. Except the light was in the pantry, not the bedroom. And her bicycle's still chained to the rack behind the library." A tray slid into the case. The rolls glistened, sugar crackling as they cooled. "You checked the rack?" "On my way here. Same lock. Same basket with the library book she never returned." Steam curled from the kettle. For a moment the only sound was the oven ticking as it cooled. "Then someone should take her the bread." "Someone should. Before the crust goes to waste." "And before whoever left that pantry light decides to come back for the purse."
created_at
"2026-09-23T09:38:27.477Z"
9/23/2026, 9:38:27 AM
model
"x-ai/grok-4.7-20260916"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
563
moderation_latency
(null)
generation_time
7301
tokens_prompt
78
tokens_completion
333
native_tokens_prompt
1304
native_tokens_completion
388
native_tokens_completion_images
(null)
native_tokens_reasoning
130
native_tokens_cached
1280
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0024128
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790156307-oVbBWh9XEXN6b55doTAA"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790156307-ljUroAJXt95hRgJmLW05"
upstream_id
"97cf0d43-8916-9462-b3ed-0f308add1cbf"
provider_responses
0
endpoint_id
"c942f37a-87e1-4fa8-83d0-8c67fda4fb4a"
id
"97cf0d43-8916-9462-b3ed-0f308add1cbf"
is_byok
false
latency
182
model_permaslug
"x-ai/grok-4.7-20260916"
provider_name
"xAI"
status
200
total_cost
0.0024128
cache_discount
0.001536
upstream_inference_cost
0
provider_name
"xAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Count dialogue tags
n/a
neededClean
false
dialogueTags
(empty)