Words
87
Reading
1 min
Listen
Play
5M
One-Shot Promoting is Bad for MoE Models?!!
This video explores why incremental prompting, one aspect of the problem at a time, is better for MoE LLMs compared to asking the model to do everything at once. Not all experts are activated for each token, so the context from coding one aspect of the program can poison the bigger picture if you don't separate requests.
I noticed this too even with big SOTA models, so it's doubly true for Local LLMs! #ai #localai #mixtureofexperts #askleo #technology
RE: LeoThread 2026-04-11 16-40