TGViewer
Skoltech Global Skoltech Global @skoltech_en · 3K subscribers
Post #2721 485
🧠 How to keep language models honest?

Large language models are great at almost everything, except one thing: they excel at fabricating convincing nonsense, even when connected to knowledge bases via RAG. Traditional methods for tackling hallucinations usually demand either tons of labeled data or heavy computations involving repeated answer generation.

Researchers from Skoltech and Sber's Center for Applied AI have found a much more elegant solution — a new method called TOHA (TOpology-based HAllucination detector).

🔹 Instead of putting the model through heavy extra workloads, the approach looks "under the hood" to analyze the topological structure of its attention maps using the MTop-Div metric. If the response graph diverges from the context, the algorithm instantly flags the discrepancy.
🔹 The method requires no training of additional classifiers and uses minimal labeled data.
🔹 In terms of accuracy, it easily competes with resource-heavy alternatives like SelfCheckGPT.

The tool is already integrated into Sber's open-source SIRIN library, ready to protect corporate knowledge bases from AI fabrications.

The research has been accepted to the main track of the A* ACL 2026 conference. Dive into the details and author comments in our full release.
  • 👍 2
More from @skoltech_en
  1. Sep 15, 2026📚 Meet the latest addition to our book column #book from a mentor of the Innovation Works…
  2. Sep 14, 2026📺 Why do we find ourselves rewatching Twilight or Friends for the hundredth time when we'…
  3. Sep 11, 2026⚛️ How can we make hydrogen-powered transport more affordable? Platinum-based catalysts ar…
  4. Sep 10, 2026🧠 Skoltech to host new Russia-China joint research and education center for neurosciences…
  5. Sep 10, 2026🎉 Skoltech leads Russian universities in A* AI publication contribution Skoltech research…
  6. Sep 9, 2026Post #2745
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →