NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36V1BJGGPBKE326JEVAZ52Y
Run Details
Model:
Grok 4.7 (Reasoning, High)
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
19.3s
Cost:
$0.005800
Provider:
xAI
Evaluation summary
Score
Evaluator
Details Link
90.48%
Matches word count
(details)
99.01%
Dialogue to Total Word Ratio
(details)
94.7471%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly slid into the booth and said, "Bill, the party starts at two. Presents sit on the porch until cake." "A porch," Bill said. "With balloons. And parents." "Parents watch the magician. We watch the table." "Those gifts belong to a kid named Trevor." "Trevor gets a duplicate from Grandma. We get the wagon." "We don't have a wagon." "We have your trunk and my excuse. You're the balloon guy. I'm the face painter who forgot her kit." "I hate balloons." "Then hold the sign that says Free Stickers and look busy. When they sing, we load the bags and leave before the clown notices." "What if the clown notices?" "Clowns notice squeaky shoes, not quiet thieves. Wear the soft ones." "And if a mom asks our names?" "You say Uncle Bill. I say Aunt Molly. We smile. We do not explain the dinosaur paper." Bill rubbed his face. "This is a terrible plan." "It's a short plan. In, grab, gone, ice cream." "If we get caught, I'm blaming you." "If we get caught, you're holding a bow the size of your head. Walk casual. Don't wave at the piƱata."
created_at
"2026-09-23T09:55:03.382Z"
9/23/2026, 9:55:03 AM
model
"x-ai/grok-4.7-20260916"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
620
moderation_latency
(null)
generation_time
19292
tokens_prompt
75
tokens_completion
427
native_tokens_prompt
1303
native_tokens_completion
1094
native_tokens_completion_images
(null)
native_tokens_reasoning
820
native_tokens_cached
1280
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0058
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790157303-oadmkGiiRjdvQBMaEv1A"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790157303-oZU0DOdexzlrciIQczG0"
upstream_id
"63d8a755-b47d-9929-b12f-547abee8df4d"
provider_responses
0
endpoint_id
"c942f37a-87e1-4fa8-83d0-8c67fda4fb4a"
id
"63d8a755-b47d-9929-b12f-547abee8df4d"
is_byok
false
latency
620
model_permaslug
"x-ai/grok-4.7-20260916"
provider_name
"xAI"
status
200
total_cost
0.0058
cache_discount
0.001536
upstream_inference_cost
0
provider_name
"xAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
90.48%
Matches word count
n/a
neededClean
false
words
190
99.01%
Dialogue to Total Word Ratio
Ratio: 93.16%, Deviation: 3.16%
neededClean
false
wordsTotal
190
wordsDialogue
177
94.7471%