Post #2189
4.79K
a new NLU benchmark for testing the ability of models to break down a question into the required steps for computing its answer. https://allenai.github.io/Break/ A work by Tomer Wolfson, accepted to TACL 2020.
AI AI, Python, Cognitive Neuroscience @ai_python_en · 3.45K subscribers