sao10k/l3.1-euryale-70b
Llama 3.1 Euryale 70B v2.2 via OpenRouter
Release Date
Aug 26th, 2024Parameters
70BContext Size
8kCreative writing
22.95%Rule following
48.04%Utility
62.89%Mathematics
75.00%Tooling
49.36%Language
63.68%Logic
71.25%Data extraction
Extract key details from a given block of text.
| Scenario | Run 1 | Run 2 | Run 3 | Run 4 | Run 5 | Run 6 | Run 7 | Run 8 | Run 9 | Run 10 | Total |
|---|---|---|---|---|---|---|---|---|---|---|---|
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 50% | 95% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 50% | 50% | 50% | 50% | 80% | |
| 100% | 100% | 100% | 100% | 100% | 50% | 50% | 50% | 50% | 50% | 75% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 50% | 50% | 90% | |
| 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 50% | 0% | 0% | 75% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 0% | 0% | 0% | 70% | |
| 100% | 50% | 50% | 50% | 50% | 50% | 50% | 50% | 50% | 50% | 55% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 50% | 50% | 50% | 85% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | |
| 77.08% | |||||||||||
Dialogue tags
Various tasks related to dialogue tags in text.
| Scenario | Run 1 | Run 2 | Run 3 | Run 4 | Run 5 | Run 6 | Run 7 | Run 8 | Run 9 | Run 10 | Total |
|---|---|---|---|---|---|---|---|---|---|---|---|
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 61% | 0% | 0% | 76% | |
| 49% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 5% | |
| 62% | 33% | 18% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 11% | |
| 53% | 51% | 50% | 4% | 2% | 0% | 0% | 0% | 0% | 0% | 16% | |
| 50% | 42% | 14% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 11% | |
| 42% | 40% | 37% | 18% | 3% | 0% | 0% | 0% | 0% | 0% | 14% | |
| 75% | 64% | 50% | 47% | 32% | 6% | 2% | 0% | 0% | 0% | 28% | |
| 22.95% | |||||||||||
Language Comprehension
Does the model understand more than just English?
| Scenario | Run 1 | Run 2 | Run 3 | Run 4 | Run 5 | Total |
|---|---|---|---|---|---|---|
| 100% | 100% | 100% | 100% | 100% | 100% | |
| 100% | 100% | 100% | 100% | 100% | 100% | |
| 100% | 0% | 0% | 0% | 0% | 20% | |
| 100% | 0% | 0% | 0% | 0% | 20% | |
| 60.00% | ||||||
Language Writing
Can the model generate text in different languages?
| Scenario | Run 1 | Run 2 | Run 3 | Run 4 | Run 5 | Total |
|---|---|---|---|---|---|---|
| 100% | 100% | 100% | 100% | 58% | 92% | |
| 100% | 85% | 67% | 64% | 0% | 63% | |
| 94% | 78% | 73% | 73% | 0% | 64% | |
| 100% | 90% | 80% | 50% | 0% | 64% | |
| 100% | 71% | 50% | 33% | 0% | 51% | |
| 66.62% | ||||||
Novel outline
Handle questions about the outline of a novel in various formats
| Scenario | Run 1 | Run 2 | Run 3 | Run 4 | Run 5 | Run 6 | Run 7 | Run 8 | Run 9 | Run 10 | Total |
|---|---|---|---|---|---|---|---|---|---|---|---|
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 0% | 90% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | |
| 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | |
| 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | |
| 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | |
| 100% | 100% | 100% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 30% | |
| 100% | 100% | 100% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 30% | |
| 100% | 100% | 100% | 100% | 0% | 0% | 0% | 0% | 0% | 0% | 40% | |
| 100% | 50% | 50% | 50% | 50% | 50% | 50% | 50% | 0% | 0% | 45% | |
| 50% | 50% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 10% | |
| 45.42% | |||||||||||
Tool usage within Novelcrafter
Output messages that are related to tool usage within Novelcrafter
| Scenario | Run 1 | Run 2 | Run 3 | Run 4 | Run 5 | Run 6 | Run 7 | Run 8 | Run 9 | Run 10 | Total |
|---|---|---|---|---|---|---|---|---|---|---|---|
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 67% | 97% | |
| 96.67% | |||||||||||
N-Length Sentences
Write sentences with exactly N words
| Scenario | Run 1 | Run 2 | Run 3 | Run 4 | Run 5 | Run 6 | Run 7 | Run 8 | Run 9 | Run 10 | Total |
|---|---|---|---|---|---|---|---|---|---|---|---|
| 100% | 100% | 100% | 100% | 93% | 89% | 86% | 84% | 76% | 62% | 89% | |
| 95% | 93% | 93% | 86% | 79% | 77% | 75% | 75% | 70% | 60% | 80% | |
| 83% | 74% | 66% | 64% | 59% | 57% | 47% | 1% | 0% | 0% | 45% | |
| 71.49% | |||||||||||
Voice/dialogue sheets
Extract dialogue from given text as voice sheets.
| Scenario | Run 1 | Run 2 | Run 3 | Run 4 | Run 5 | Run 6 | Run 7 | Run 8 | Run 9 | Run 10 | Total |
|---|---|---|---|---|---|---|---|---|---|---|---|
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 0% | 0% | 80% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 0% | 0% | 0% | 70% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 0% | 90% | |
| 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 0% | 0% | 80% | |
| 64.00% | |||||||||||
Write N of X
Write exactly N words/sentences/paragraphs...
| Scenario | Run 1 | Run 2 | Run 3 | Run 4 | Run 5 | Run 6 | Run 7 | Run 8 | Run 9 | Run 10 | Total |
|---|---|---|---|---|---|---|---|---|---|---|---|
| 100% | 100% | 100% | 100% | 100% | 100% | 98% | 98% | 77% | 77% | 95% | |
| 100% | 98% | 92% | 92% | 27% | 9% | 2% | 0% | 0% | 0% | 42% | |
| 54% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 5% | |
| 54% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 5% | |
| 92% | 54% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 15% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | |
| 100% | 100% | 100% | 98% | 98% | 92% | 77% | 77% | 77% | 0% | 82% | |
| 100% | 100% | 100% | 98% | 54% | 0% | 0% | 0% | 0% | 0% | 45% | |
| 100% | 100% | 100% | 0% | 0% | 0% | 0% | 0% | 0% | 0% | 30% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | |
| 100% | 100% | 100% | 100% | 100% | 100% | 100% | 0% | 0% | 0% | 70% | |
| 100% | 100% | 100% | 100% | 0% | 0% | 0% | 0% | 0% | 0% | 40% | |
| 56.14% | |||||||||||