NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYCTNF6SB27WDT36X9GMW84M
Run Details
Model:
Gemini 3.5 Flash Lite (Reasoning)
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
1.8s
Cost:
$0.000613
Provider:
Google AI Studio
Evaluation summary
Score
Evaluator
Details Link
60.28%
Matches word count
(details)
33.66%
Dialogue to Total Word Ratio
(details)
46.9680%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly checked her wristwatch while adjusting her ridiculously oversized clown shoes. Bill struggled manfully to squeeze his broad shoulders into a sequined fairy princess vest that smelled faintly of strawberry vape juice. They crouched behind a strategically placed plastic flamingo in the overgrown bushes of suburban cul de sac hell. "Are you sure this is the target, Bill? The intelligence was sketchy," Molly whispered, glaring at a rogue piece of cotton candy stuck to her velvet sleeve. "Trust me, the intel is solid. Little Timmy’s uncle is a crypto bro who bought out the entire catalog of a luxury toy store. We are talking solid gold Hot Wheels and limited-edition action figures," Bill replied, popping a breath mint into his mouth. "And you are certain the security guard is neutralized?" "Uncle Gary passed out after three glasses of cheap chardonnay. He is currently snoring louder than a lawnmower." "Then let us infiltrate the bouncy castle of doom. Remember, grab the big boxes first, and if anyone asks, we are the hired entertainment." "I am not doing balloon animals, Molly. I draw the line at dachshunds."
created_at
"2026-07-25T14:25:31.362Z"
7/25/2026, 2:25:31 PM
model
"google/gemini-3.5-flash-lite-20260721"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
289
moderation_latency
(null)
generation_time
1739
tokens_prompt
75
tokens_completion
288
native_tokens_prompt
67
native_tokens_completion
237
native_tokens_completion_images
0
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"STOP"
service_tier
"default"
usage
0.0006126
router
(null)
provider_responses
0
endpoint_id
"6bd8f433-79e1-416c-b407-1772eb796c9a"
id
"W8dkasvrGbLVjMcPx5-fqQY"
is_byok
false
latency
289
model_permaslug
"google/gemini-3.5-flash-lite-20260721"
provider_name
"Google AI Studio"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1784989531-VTgFK010VPxrd8hHBEO7"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1784989531-31GwdO88gZcb0pHaJW0A"
upstream_id
"W8dkasvrGbLVjMcPx5-fqQY"
total_cost
0.0006126
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Google AI Studio"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
60.28%
Matches word count
n/a
neededClean
false
words
185
33.66%
Dialogue to Total Word Ratio
Ratio: 60.22%, Deviation: 10.22%
neededClean
false
wordsTotal
186
wordsDialogue
112
46.9680%