NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36VPY1464Y6HPPD7M7XAX8C
Run Details
Model:
Grok 4.7 (Reasoning, High)
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
23.9s
Cost:
$0.006990
Provider:
xAI
Evaluation summary
Score
Evaluator
Details Link
99.98%
Matches word count
(details)
93.71%
Dialogue to Total Word Ratio
(details)
96.8448%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly checked the driveway. "Bill, it is a children's party. There is a clown. There is a pony. We are not stealing the Declaration of Independence." "Presents are presents," Bill said. "The table is by the fence. Everyone stares at the cake. We walk in, we walk out." "Walk in dressed as what, disappointed uncles?" "We already look like disappointed uncles." "That is not a plan. That is a cry for help." "One box is huge. Shiny paper. I bet it is something good." "You bet. With our bail money." "Molly." "Fine. You take the left side of the gift table. I stand near the clown and ask him where the bathroom is. Clowns hate questions." "And if a kid sees us?" "We say we are the bad magicians. Then we leave before anyone asks for a trick." "No tricks." "Especially not smart ones. Grab, waddle, van. If the pony looks at you, you lose." "I am not afraid of a pony." "You should be. It has better posture than our whole operation." Bill sighed. "So we are doing this." "We are doing the stupid version," Molly said. "Quiet feet. No speeches. And if the cake comes out, we are already gone."
created_at
"2026-09-23T10:06:50.412Z"
9/23/2026, 10:06:50 AM
model
"x-ai/grok-4.7-20260916"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
854
moderation_latency
(null)
generation_time
23832
tokens_prompt
75
tokens_completion
530
native_tokens_prompt
1303
native_tokens_completion
1342
native_tokens_completion_images
(null)
native_tokens_reasoning
1050
native_tokens_cached
1280
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0069904
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790158010-K7noNCchQZlejNA8A0bZ"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790158010-xBx1YZhoZXxdtzls9AGd"
upstream_id
"7522dfc7-64bf-9711-9095-c88bb40d29ed"
provider_responses
0
endpoint_id
"c942f37a-87e1-4fa8-83d0-8c67fda4fb4a"
id
"7522dfc7-64bf-9711-9095-c88bb40d29ed"
is_byok
false
latency
96
model_permaslug
"x-ai/grok-4.7-20260916"
provider_name
"xAI"
status
200
total_cost
0.0069904
cache_discount
0.001536
upstream_inference_cost
0
provider_name
"xAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
99.98%
Matches word count
n/a
neededClean
false
words
202
93.71%
Dialogue to Total Word Ratio
Ratio: 95.05%, Deviation: 5.05%
neededClean
false
wordsTotal
202
wordsDialogue
192
96.8448%