Tired of LLMs forgetting everything after 20 turns?
I built Acervo to fix that: a graph-based memory layer that keeps token usage almost constant, even in 100+ turn conversations.
After 360+ turns of testing… the results surprised even me.
Thread on what v0.3 actually delivers ↓
El futbolero está en su casa, masticando bronca, llorando, tratando de retomar fuerzas para encarar uno de los lunes más duros de su vida. Acá están los que ven futbol cada cuatro años y mucho no entienden, a veces me gustaría ser uno de ellos.
Kimi K3 is the best performing model on https://t.co/aporqgIfIh, ahead of Fable, reaching a comparable success rate in less time.
This is the first time that an open model is ahead of all proprietary ones for this comprehensive web engineering benchmark.
Notes:
▪️ Benchmarks don’t always tell the full story, although this is important signal, adding to mounting evidence that this could be a breakthrough moment for open models
▪️ No model as of yet has reached 100% completion on this set of evals. The top performer peaks at 92% and 96% “with help”
Bajo un intenso frío, Tricao Malal salió otra vez a festejar el triunfo de Argentina🇦🇷. El pueblo del norte neuquino, rodeado del majestuoso Volcán Domuyo, el cerro El Palao y el Volcán Trómen, mantiene la ilusión intacta, como cada hincha de cualquier punto del país...💙🤍💙