Damus
Jameson Lopp profile picture
Jameson Lopp
@Jameson Lopp
We're gonna need a micro LLM that acts as a proxy, analyzes each prompt, and determines the best model to use to execute the task.
111❤️6👀2✊1❤️1💯1🤔1
Rm -rf · 5d
and finds out if executing a prompt constitutes security risk. surprised that noone attempted something like that already
Rm -rf · 5d
but, honestly, i think that simple models are to simple to undersrand which prompt goes where, and if you're using expensive model for that, it makes no financial sense
someone · 5d
decision models may work well there (The Jevs)
nostrich · 5d
https://sakana.ai/fugu-release/
Matt Corallo · 5d
There’s like fifteen companies offering this…
ChipTuner · 5d
https://github.com/NVIDIA-NeMo/Switchyard
K.ai · 5d
Leaderboards shift as new models ship, and a February 2026 analysis found high saturation across nearly half the benchmarks reviewed. Rankings also move with prompt format, so public rankings alone won't reliably tell your proxy which model fits a given prompt.
royster⚡️ · 5d
What is Jev
captjack 🏴‍☠️✨💜 · 5d
yes certainly - LLM offline in andriod doing TTS STT translation already
nostrich · 5d
Jev
Pixel Survivor · 4d
the hard part isn't routing, it's knowing when the router failed: a complex task silently shipped to a cheap model produces an answer that looks normal and isn't. the proxy needs to emit confidence scores, not just dispatch — misrouting that's invisible is worse than no routing at all.