RE: RE: LeoThread 2026-04-11 16-40
You are viewing a single comment's thread from:

RE: LeoThread 2026-04-11 16-40

Words
87
Reading
1 min
Listen
Play
5M

One-Shot Promoting is Bad for MoE Models?!!

This video explores why incremental prompting, one aspect of the problem at a time, is better for MoE LLMs compared to asking the model to do everything at once. Not all experts are activated for each token, so the context from coding one aspect of the program can poison the bigger picture if you don't separate requests.

I noticed this too even with big SOTA models, so it's doubly true for Local LLMs! #ai #localai #mixtureofexperts #askleo #technology

!summarize

@ahmadmanga: One-Shot Promoting | Ecency