NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36TQKKPFNSGVYD867MEWTX8
Run Details
Model:
Grok 4.7 (Reasoning, High)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
1m 12s
Cost:
$0.030126
Provider:
xAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly crouched behind the hedge and treated the birthday party like a vault with frosting. Bill crouched beside her, sunglasses crooked, watching a piƱata swing over the lawn while children screamed with professional enthusiasm. Paper streamers snapped hard in the wind. A teenager guarded the gift table by scrolling on a phone and ignoring everything worth taking. Somewhere a clown practiced sword fights against balloon animals and lost. "Presents first, before the clown notices," Molly whispered. Bill nodded as if this were wisdom. "Then we take cake." "No cake, Bill." "A little cake never hurt a heist." Molly refused to dignify dessert with strategy. She pointed at the bright heap of boxes, then at their reusable shopping bag, which suggested groceries more than crime. Bill invented excuses and forgot them immediately. The dog had developed opinions about his shoes. Neither thief had brought a gift, a child, or a reason anyone would invite them. They had brought ambition and the belief that wrapping paper was unguarded treasure. A mother near the punch bowl glanced over once. Molly smiled like she belonged at the party. Bill smiled like he had misplaced a nephew he had never met, and the presents glittered.
created_at
"2026-09-23T09:49:43.931Z"
9/23/2026, 9:49:43 AM
model
"x-ai/grok-4.7-20260916"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
609
moderation_latency
(null)
generation_time
71669
tokens_prompt
75
tokens_completion
1026
native_tokens_prompt
1303
native_tokens_completion
6130
native_tokens_completion_images
(null)
native_tokens_reasoning
5885
native_tokens_cached
1152
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0301264
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790156983-Spzn1a4XOxoKka5sacvg"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790156983-NQs3o1HHrscs7KmGEPtG"
upstream_id
"376723c4-a3d2-99b9-9be4-bc6fd96e8b38"
provider_responses
0
endpoint_id
"c942f37a-87e1-4fa8-83d0-8c67fda4fb4a"
id
"376723c4-a3d2-99b9-9be4-bc6fd96e8b38"
is_byok
false
latency
572
model_permaslug
"x-ai/grok-4.7-20260916"
provider_name
"xAI"
status
200
total_cost
0.0301264
cache_discount
0.0013824
upstream_inference_cost
0
provider_name
"xAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 10.00%, Deviation: 0.00%
neededClean
false
wordsTotal
200
wordsDialogue
20
100.0000%