EXTRAI, part 3: I built a meta-analysis pipeline out of local AIs auditing each other — and the quality gate did more damage than the sabotage it was https://t.co/HtEdjmgDkG via @LinkedIn
EXTRAI, part 2: no model got a single confidence interval right in its head — then I handed them a calculator, and one closed a perfect meta-analysis https://t.co/eRhJoZuqeW via @LinkedIn
EXTRAI, part 1: I made four local models redo a meta-analysis's data extraction — they found more errors in it than it found in them https://t.co/TDI4TmNp3d via @LinkedIn
Summarization benchmark, part 13: the most faithful model I've ever tested — and it lost to both smaller ones https://t.co/cNjDODTOl9 via @LinkedIn
Summarization benchmark, the wrap-up: what 432 graded measurements taught us — and where a local model can be trusted https://t.co/0u0XuvGmLx via @LinkedIn
Summarization benchmark, the wrap-up: what 432 graded measurements taught us — and where a local model can be trusted https://t.co/fgy70CNFzn via @LinkedIn
Summarization benchmark, part 11: I tried to replace the human reviewer with software — and what I got was a triage tool https://t.co/J9kqch1gD1 via @LinkedIn
Summarization benchmark, part 9: the study note — the format local models get most right (and the gap that became a hallucination) https://t.co/I8AVuFQ5Ra via @LinkedIn
Summarization benchmark, part 7: is the collapse the model's fault or Ollama's? I swapped the engine to find out https://t.co/H7zkHjTsch via @LinkedIn
Visited Sabin Labs Central Processing center! Sabin truly represents the success of the #women#entrepreneurs of Brazil! Success built on innovation, dedication to customers & commitment to quality! Parabens to Team Sabin! #healthcare#healthtech https://t.co/iCeH6nJK15