Close Menu
Techora News HubTechora News Hub
    Facebook X (Twitter) Instagram
    Techora News HubTechora News Hub
    • Home
    • Crypto News
      • Bitcoin
      • Ethereum
      • Altcoins
      • Blockchain
      • DeFi
    • AI News
    • Stock News
    • Learn
      • AI for Beginners
      • AI Tips
      • Make Money with AI
    • Reviews
    • Tools
      • Best AI Tools
      • Crypto Market Cap List
      • Stock Market Overview
      • Market Heatmap
    • Contact
    Techora News HubTechora News Hub
    Home»AI News»NVIDIA Releases Kumo Tabular: Open Tabular Foundation Models That Predict New Rows in a Single Forward Pass
    AI News

    NVIDIA Releases Kumo Tabular: Open Tabular Foundation Models That Predict New Rows in a Single Forward Pass

    October 1, 2026
    Facebook Twitter Pinterest Telegram LinkedIn Tumblr WhatsApp Email
    NVIDIA Releases Kumo Tabular: Open Tabular Foundation Models That Predict New Rows in a Single Forward Pass
    Share
    Facebook Twitter LinkedIn Pinterest Telegram Email
    ledger


    NVIDIA has released Kumo Tabular, a new family of tabular foundation models (TFMs) for classification and regression. If you have followed TabPFN or TabICL, the setup will look familiar. The model takes labeled rows as context and predicts new rows in one forward pass. There is no training, no hyperparameter tuning, and no feature engineering.

    Kumo Tabular comes in Small, Medium, and Large versions, spanning about 28M to 215M parameters. It runs through NVIDIA’s open-source structured-data-models (SDM) library.

    Is it deployable? Yes. Weights ship under the OpenMDW-1.1 license, which permits commercial use. The SDM code is Apache-2.0, and it needs Python 3.11+ and PyTorch 2.7+, with examples targeting a CUDA GPU.

    What the SDM Library Adds

    SDM is a GPU-native library for structured-data foundation models and preprocessing. Besides Kumo Tabular, it ships TabICLv2, Google’s TabFM, and KumoRelational for multi-table data. All models share one in-context learning interface built on a TableTensor container. The library also handles preprocessing, ensembling, and many-class prediction.

    murf

    How Kumo Tabular Works

    Kumo Tabular is a Transformer built around the structure of a table. It uses column, row, and in-context attention, as introduced in TabICL and TabPFN. The pipeline has 3 stages:

    • Cell embedding: Numerical and categorical values pass through learned Fourier features, with separate weights per type. Missing values need no imputation.
    • Row embedding: Column attention uses induced self-attention, so cost grows linearly with rows. Row attention, with rotary positions, learns feature interactions. 4 learnable [CLS] tokens compress each row.
    • In-context learning: A final Transformer runs over row embeddings. Context rows attend to each other, while query rows attend only to context rows.

    Because the context never sees the queries, its keys and values are computed once and reused. The head outputs class probabilities, or 999 quantiles for regression. That gives a point prediction plus an uncertainty estimate.

    One more detail matters at scale. Softmax attention spreads thin as the number of keys grows. Kumo Tabular scales each query by a temperature that grows with the log of the key count. The coefficient is learned per attention head, so attention stays sharp on larger tables.

    Trained Only on Artificial Tables

    Kumo Tabular is pretrained entirely on synthetic tables sampled from Structural Causal Models (SCMs). A random causal graph links hidden variables through linear maps, small neural networks, trees, or Gaussian processes. The generator also injects messy, real-world patterns: missing values, high-cardinality categories, heavy-tailed targets, and conflicting duplicate rows.

    Training ran in 3 stages, similar to TabICLv2. Context grew from 1,024 rows to 60,000 rows, with up to 100 columns. Small, Medium, and Large saw about 35M, 71M, and 137M artificial tables. Classification and regression are trained as separate models. NVIDIA says the training recipe and data generators will be released soon.

    Benchmarks

    With default settings, Kumo Tabular ranks first overall on TabArena with an Elo of 1950. NVIDIA team reports it runs 17x faster than LimiX-2 on a single RTX 6000 Pro. All 3 sizes sit on the accuracy and inference-time Pareto front.

    • BeyondArena: First place, with an Elo of 1418 and an Improvability score of 7.78%.
    • TALENT: Top overall ranking, with average ranks of 6.67 (accuracy), 3.98 (log-loss), and 4.22 (RMSE).
    • ScoringBench: Large and Medium rank first and second on average rank.

    Kumo Tabular vs Its Closest Competitors

    FeatureKumo TabularTabICLv2TabPFN-3LimiX-2TabFMDeveloperNVIDIAInria SODAPrior LabsStable AIGoogle ResearchParameters~28M to 215M (3 sizes)27.55M (cls), 28.54M (reg)Not listed in docs400M~1.64BTasksClassification, regressionClassification, regressionClassification, regressionClassification, regression, imputationClassification, regressionNative classes per pass10 (ECOC for more)10 (hierarchical for more)160Not specified10 (hard limit)Weights licenseOpenMDW-1.1BSD-3-ClauseTABPFN-3 License v1.0StableAI LimiX Non-CommercialTabFM Non-Commercial v1.0Commercial use of weightsYesYesPaid license requiredNoNoRuns in NVIDIA SDMYesYesNoNoYes

    Sources: NVIDIA blog, SDM model docs, Prior Labs docs, LimiX GitHub, TabICLv2 paper. Checked September 30, 2026.

    The license row is the real differentiator. TabPFN-3, LimiX-2, and TabFM weights carry non-commercial terms. Kumo Tabular and TabICLv2 are the permissive options, and Kumo Tabular leads the benchmarks NVIDIA reports.

    Getting Started

    Install the library and pass a DataFrame through TableTensor, following the model card:

    # pip install structured-data-models
    from sklearn.datasets import load_breast_cancer
    import sdm

    df = load_breast_cancer(as_frame=True).frame
    table = sdm.TableTensor.from_pandas(
    df=df,
    stypes=sdm.infer_stypes(df, overrides={“target”: “categorical”}),
    device=”cuda”,
    )
    model = sdm.models.KumoTabular(task=”classification”, device=”cuda”)
    probs = model(
    x_context=table[:300].drop_columns(“target”),
    y_context=table[:300, “target”],
    x_query=table[300:].drop_columns(“target”),
    num_estimators=8,
    )

    The size argument accepts “small”, “medium”, or “large”, and defaults to large.

    Key Takeaways

    • NVIDIA’s Kumo Tabular predicts new table rows in 1 forward pass, with no training.
    • 3 sizes span about 28M to 215M parameters, pretrained only on synthetic tables.
    • NVIDIA reports first place on TabArena (Elo 1950), BeyondArena, TALENT, and ScoringBench.
    • OpenMDW-1.1 weights allow commercial use, unlike TabPFN-3, LimiX-2, and TabFM.
    • It runs through NVIDIA’s GPU-native SDM library alongside TabICLv2, TabFM, and KumoRelational.

    Check out the Model on HF, GitHub Repo and Technical details. All credit goes to the researcher of this project. Also, feel free to follow us on Twitter and don’t forget to join our 150k+ML SubReddit and Subscribe to our Newsletter. Wait! are you on telegram? now you can join us on telegram as well.

    Need to partner with us for promoting your GitHub Repo OR Hugging Face Page OR Product Release OR Webinar etc.? Connect with us

    Asif Razzaq is the CEO of Marktechpost AI Media Inc.. As a visionary entrepreneur and engineer, Asif is committed to harnessing the potential of Artificial Intelligence for social good. His most recent endeavor is the launch of an Artificial Intelligence Media Platform, Marktechpost, which stands out for its in-depth coverage of machine learning and deep learning news that is both technically sound and easily understandable by a wide audience. The platform boasts of over 2 million monthly views, illustrating its popularity among audiences.



    Source link

    quillbot
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email

    Related Posts

    New tool lets users repair AI-generated 3D models, then fabricate them just the way they want | MIT News

    October 6, 2026

    Documenting the tech worker movement | MIT News

    October 5, 2026

    Computational tools for society’s most complex challenges | MIT News

    October 4, 2026

    Meta, OpenAI and Uber Just Taught AI Agents to Talk First. What About When to Stay Quiet?

    October 3, 2026

    3 Questions: A new resource to empower young entrepreneurs | MIT News

    October 2, 2026

    Who we become when we talk to machines | MIT News

    September 30, 2026
    kraken
    Latest Posts

    New tool lets users repair AI-generated 3D models, then fabricate them just the way they want | MIT News

    October 6, 2026

    Trump’s $5,000 Checks Could Send Billions Into Bitcoin and Crypto: But There’s a Catch

    October 6, 2026

    How to Earn Money with AI – LAZIEST Way to Make Money With AI – AI Remote Jobs – AI Videos Money

    October 6, 2026

    AI Agents Explained for Complete Beginners (START HERE)

    October 6, 2026

    AI hacks now operating at machine speed, not hacker speed, says Fmr. White House CIO Theresa Payton

    October 6, 2026
    10web
    LEGAL INFORMATION
    • Privacy Policy
    • Terms Of Service
    • Social Media Disclaimer
    • DMCA Compliance
    • Anti-Spam Policy
    Top Insights

    Solana DvP settlement requires 100% upfront cash for every trade

    October 7, 2026

    Ethereum’s Glamsterdam Upgrade Launches on Sepolia Testnet

    October 7, 2026
    kraken
    Facebook X (Twitter) Instagram Pinterest
    © 2026 TechoraNewsHub.com - All rights reserved.

    Type above and press Enter to search. Press Esc to cancel.

    bitcoin
    Bitcoin (BTC) $ 84,155.00
    ethereum
    Ethereum (ETH) $ 2,613.51
    tether
    Tether (USDT) $ 0.999906
    bnb
    BNB (BNB) $ 767.44
    xrp
    XRP (XRP) $ 1.47
    usd-coin
    USDC (USDC) $ 0.999962
    solana
    Solana (SOL) $ 118.38
    tron
    TRON (TRX) $ 0.332419
    staked-ether
    Lido Staked Ether (STETH) $ 2,265.05
    figure-heloc
    Figure Heloc (FIGR_HELOC) $ 1.04