NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYCSM30BXA4NKJ79ZPHABRWH
Run Details
Model:
Gemini 3.5 Flash Lite (Reasoning, Minimal)
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
1.8s
Cost:
$0.000715
Provider:
Google AI Studio
Evaluation summary
Score
Evaluator
Details Link
99.38%
Matches word count
(details)
0.00%
Dialogue to Total Word Ratio
(details)
49.6885%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly glared at the giant pink velvet bow taped to the windshield of their getaway van. Bill, are you sure this is the most lucrative target in the tri-state area? Absolutely, Molly, whispered Bill, adjusting his tiny burglar beanie over his ears. Do you know how much pristine, unblemished plastic is inside those gift bags? It is a seven-year-old birthday party, Bill. Exactly. High-end Lego sets, remote-controlled helicopters, and at least three artisanal cupcakes with buttercream frosting. We are master criminals, not toddlers, Molly hissed, poking him in the ribs. We knocked off a Swiss bank last month. And we got away with eighty thousand Swiss francs in heavy, untraceable coin change that strained my lower back, Bill retorted. This is ergonomic theft. Lightweight cardboard boxes. Minimal police response. What about the security guard? His name is Barnaby, he is six feet tall, and he is currently wearing a paper pirate hat and eating a juice box. Fine. But if I get attacked by a horde of sugar-crazed children wielding juice boxes, I am keeping the remote-controlled helicopter for myself. Deal. Now sync your walkie-talkie to my SpongeBob frequency and let's go steal some juice.
created_at
"2026-07-25T14:07:17.525Z"
7/25/2026, 2:07:17 PM
model
"google/gemini-3.5-flash-lite-20260721"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
230
moderation_latency
(null)
generation_time
1771
tokens_prompt
75
tokens_completion
310
native_tokens_prompt
67
native_tokens_completion
278
native_tokens_completion_images
0
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"STOP"
service_tier
"default"
usage
0.0007151
router
(null)
provider_responses
0
endpoint_id
"6bd8f433-79e1-416c-b407-1772eb796c9a"
id
"FcNkavLWI-yW-8YP_ai40Qg"
is_byok
false
latency
230
model_permaslug
"google/gemini-3.5-flash-lite-20260721"
provider_name
"Google AI Studio"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1784988437-zkmLhiF1VcozLBfmXn3i"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1784988437-bHVSgTTZVPuGJKjVPaM6"
upstream_id
"FcNkavLWI-yW-8YP_ai40Qg"
total_cost
0.0007151
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Google AI Studio"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
99.38%
Matches word count
n/a
neededClean
false
words
195
0.00%
Dialogue to Total Word Ratio
Ratio: 0.00%, Deviation: 90.00%
neededClean
false
wordsTotal
202
wordsDialogue
0
49.6885%