Do tabular foundation models repair themselves?

·LessWrong··

TL;DR Ablation is a widely adopted technique in mechanistic interpretability research. Self-repair invalidates conclusions drawn from ablation experiments. While self-repair is prevalent in language models, we ask whether it exists in tabular foundations models (TFMs). Our experiments on 4 SOTA tabular foundation models and 15 binary classification tasks suggest a negative result. Redundancy is more likely instead. The full version of our work is available at https://xiaohan2012.github.io/articl...

Read full article →

Related Articles

Hackers Got Inside a Flock Camera
driverdan · Hacker News · 15h ago
Apple Reference Image: A New Approach for Verified Photography
imwally · Hacker News · 1d ago
Xiaomi Mimo 2.6 live post-training dashboard
krackers · Hacker News · 8h ago
Training a 4B model to produce 81% faster query plans than Postgres
polyphilz · Hacker News · 9h ago
Nvidia announces native GPU programming in Rust
nonmaskable · Hacker News · 17h ago