NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36WHESY2G0RP18KW8FRBNYX
Run Details
Model:
Grok 4.7 (Reasoning, High)
Test:
Dialogue tags
Scenario:
Write unattributed dialogue
Duration:
15.8s
Cost:
$0.004659
Provider:
xAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Count dialogue tags
(details)
Model input
User/Human
You are an expert novel writer in the cozy mystery genre. Write a scene of about 200 words, with dialogue between two characters. The dialogue should be unattributed, meaning the speaker is not explicitly mentioned. The scene should be set in a small town bakery. Only return the story, nothing else.
Model response
Messages
Meta
AI
The morning rush had thinned to crumbs and the tick of the cooling oven. Cinnamon hung in the air, sweet and a little sharp. “She was in here before dawn. Wouldn’t meet my eye.” “Mrs. Pell never misses the first tray. Not in twenty years.” “Asked for the usual, then changed it. Said the usual didn’t sit right anymore.” A knife drew a slow line through frosting. Outside, the square was still blue with cold. “And the ribbon?” “Gone from the window. She noticed before I did. Stood there counting the empty hook like it might confess.” “Sheriff called it a prank.” “Sheriff didn’t see the latch. Wiped. Not dusty-wiped—careful. And a print in the butter, small, like someone leaned in to listen.” The bell gave a tired chime as wind pushed the door. Neither of them looked up. “So we keep the spare key where it is.” “We keep it. And we watch who asks whether the cinnamon rolls sold out early.” “If they already know the answer?” “Then we’ll know who moved the ribbon.” Flour drifted in a slant of light. Somewhere in the back, a timer began to ring, soft and patient, as if the morning still had time to tell the truth.
created_at
"2026-09-23T10:21:19.559Z"
9/23/2026, 10:21:19 AM
model
"x-ai/grok-4.7-20260916"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
496
moderation_latency
(null)
generation_time
15719
tokens_prompt
78
tokens_completion
416
native_tokens_prompt
1304
native_tokens_completion
824
native_tokens_completion_images
(null)
native_tokens_reasoning
547
native_tokens_cached
1152
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0046592
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790158879-QAYrvuYMZFWu3bPKnZqz"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790158879-xf2o3zj4YDC3wunET2gl"
upstream_id
"c82b2b43-3f53-9579-aca2-2d9ec5714948"
provider_responses
0
endpoint_id
"c942f37a-87e1-4fa8-83d0-8c67fda4fb4a"
id
"c82b2b43-3f53-9579-aca2-2d9ec5714948"
is_byok
false
latency
379
model_permaslug
"x-ai/grok-4.7-20260916"
provider_name
"xAI"
status
200
total_cost
0.0046592
cache_discount
0.0013824
upstream_inference_cost
0
provider_name
"xAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Count dialogue tags
n/a
neededClean
false
dialogueTags
(empty)