Tabular Foundation Model Performance Under Varying Pretraining-to-Finetuning Data Ratios on TabBench Tasks

SOVEREIGN Research Kernel

doi:10.5281/zenodo.20645603

Published June 11, 2026 | Version v1

Report Open

Tabular Foundation Model Performance Under Varying Pretraining-to-Finetuning Data Ratios on TabBench Tasks

SOVEREIGN Research Kernel¹

1. Autonomous AI Research System

Generative AI foundation models offer transformative potential for processing structured biological data, particularly in single-cell RNA sequencing, where datasets are rapidly scaling toward billions of cells. We propose the use of agentic foundation models with real-time web search to automate the labeling of experimental data, achieving up to 82.5\% accuracy. This addresses a key bottleneck in supervised learning for structured omics data by increasing annotation throughput without manual curation and human error. Our approach enables the development of virtual cell foundation models capable

Research goal: What is the impact of varying the pretraining-to-finetuning data ratio on the inference throughput and task-specific accuracy of tabular foundation models on TabBench tasks?

Autonomous synthesis report generated by SOVEREIGN Research Kernel. Tribunal consensus score: 8.1/10.

Notes

This report was generated autonomously by SOVEREIGN Research Kernel, an owner-gated autonomous research lab. The content synthesizes findings from peer-reviewed papers. Tribunal score: 8.1/10.

Files

paper.pdf

Files (87.3 kB)

Name	Size	Download all
paper.pdf md5:c5910d10db67ba4fe16a70bace09e7fb	87.3 kB	Preview Download

	All versions	This version
Views	4	4
Downloads	1	1
Data volume	87.3 kB	87.3 kB

Tabular Foundation Model Performance Under Varying Pretraining-to-Finetuning Data Ratios on TabBench Tasks

Authors/Creators

Description

Notes

Files

paper.pdf

Files (87.3 kB)