NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYCX7RW4ZF639ZX70V0R3REN
Run Details
Model:
Thinking Machines Inkling (Reasoning)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
13m 37s
Cost:
$0.040658
Provider:
DeepInfra
Evaluation summary
Score
Evaluator
Details Link
97.63%
Matches word count
(details)
99.90%
Dialogue to Total Word Ratio
(details)
98.7619%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly unfurled the blueprint across the dashboard with the flair of a magician revealing a disappointing card. The target was Timmy's birthday party, hosted in a suburban fortress of inflatable arches and cheerful streamers. Bill adjusted his black turtleneck, insisting it made him invisible despite resembling a shadow from a catalog. "We're robbing a toddler party?" Bill said, squinting. "This is rock bottom." Molly tapped the circled present table. "There are forty-seven gifts. The paper alone is gold." Bill sighed, polishing a lockpick. The plan required posing as entertainment—a magician and his suspicious rabbit. Molly had procured a top hat smelling of mothballs. Bill refused the costume but agreed to carry the sack. They synchronized watches with astronaut gravity. The window was forty-five minutes, ending when a clown twisted balloons into questionable shapes. Molly loaded glitter bombs, a decoy teddy, and earplugs. Bill warmed the getaway minivan—stolen from a librarian—while muttering about standards. "Not monsters," Molly said, sliding behind the wheel. "Just adaptable." The engine coughed. Outside, the house blazed with lights like a beacon of poor judgment. They drove slowly, committing with the doomed enthusiasm that defined their partnership.
created_at
"2026-07-25T15:10:28.238Z"
7/25/2026, 3:10:28 PM
model
"thinkingmachines/inkling-20260715"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
901
moderation_latency
(null)
generation_time
816623
tokens_prompt
75
tokens_completion
8999
native_tokens_prompt
75
native_tokens_completion
10027
native_tokens_completion_images
(null)
native_tokens_reasoning
8673
native_tokens_cached
32
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.04065779
router
(null)
provider_responses
0
endpoint_id
"d6b9d68e-115a-49d7-81cd-1ad1c6942b0e"
id
"chatcmpl-RUJO3ENWo7CEdM9WJjSkhkH7"
is_byok
false
latency
66
model_permaslug
"thinkingmachines/inkling-20260715"
provider_name
"DeepInfra"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1784992228-NgzyvGOqJpaWzxFwKQ5i"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1784992228-w2BCYQQ0XJBCh8zOQ3C4"
upstream_id
"chatcmpl-RUJO3ENWo7CEdM9WJjSkhkH7"
total_cost
0.04065779
cache_discount
0.00002656
upstream_inference_cost
0
provider_name
"DeepInfra"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
97.63%
Matches word count
n/a
neededClean
false
words
193
99.90%
Dialogue to Total Word Ratio
Ratio: 11.79%, Deviation: 1.79%
neededClean
false
wordsTotal
195
wordsDialogue
23
98.7619%