1 1 51

Max Current

eldogbbhed

AI & ML interests

None yet

Recent Activity

liked a model about 1 month ago

MeissonFlow/Meissonic

liked a model about 1 month ago

SWivid/F5-TTS

liked a model about 1 month ago

Mi6paulino/Michaelpaulino

View all activity

Organizations

None yet

eldogbbhed's activity

liked 3 models about 1 month ago

liked a Space 2 months ago

Running

🤖

Sponsorblock ML

Reacted to Smooke's post with 🧠 2 months ago

Post

598

Chomsky predicting LLMs in 1956, curated by Ryan Rhodes (Rutgers)

liked 6 models 2 months ago

FPHam/L3-8B-Everything-COT

Text Generation • Updated Jul 3 • 261 • 12

kyutai/moshiko-candle-bf16

Updated Sep 18 • 902 • 8

kyutai/moshika-pytorch-bf16

Updated Sep 18 • 541 • 45

Exthalpy/state-0

Text Generation • Updated Sep 20 • 2 • 12

ICTNLP/Llama-3.1-8B-Omni

Updated 12 days ago • 5.11k • 379

arcee-ai/Llama-3.1-SuperNova-Lite

Text Generation • Updated Oct 2 • 9.93k • 174

liked 3 models 3 months ago

SG161222/RealFlux_1.0b_Schnell

Updated Oct 8 • 76

openbmb/MiniCPM3-4B

Text Generation • Updated 12 days ago • 29.1k • 384

allenai/OLMoE-1B-7B-0924

Text Generation • Updated Oct 19 • 8.78k • 109

Reacted to mlabonne's post with 👍 3 months ago

Post

15891

Large models are surprisingly bad storytellers.

I asked 8 LLMs to "Tell me a bedtime story about bears and waffles."

Claude 3.5 Sonnet and GPT-4o gave me the worst stories: no conflict, no moral, zero creativity.

In contrast, smaller models were quite creative and wrote stories involving talking waffle trees and bears ostracized for their love of waffles.

Here you can see a comparison between Claude 3.5 Sonnet and NeuralDaredevil-8B-abliterated. They both start with a family of bears but quickly diverge in terms of personality, conflict, etc.

I mapped it to the hero's journey to have some kind of framework. Prompt engineering can definitely help here, but it's still disappointing that the larger models don't create better stories right off the bat.

Do you know why smaller models outperform the frontier models here?