Technophile NewsTechnophile News
  • Home
  • News
  • PC
  • Phones
  • Android
  • Gadgets
  • Games
  • Guides
  • Accessories
  • Reviews
  • Spotlight
  • More
    • Artificial Intelligence
    • Web Stories
    • Press Release
What's On

Minnesota Shooting Suspect Allegedly Used Data Broker Sites to Find Targets’ Addresses

16 June 2025

The Definitive Story of Tesla Takedown

16 June 2025

iQOO Z10 Lite 5G: Launch Date, Expected Price in India, Specifications, Features and More

16 June 2025

Try This Free Version of Microsoft Office That Runs in Your Browser

16 June 2025

‘Psyop’: How Far-Right Conspiracy Theories About the Minnesota Shooting Evolved to Protect MAGA

16 June 2025
Facebook X (Twitter) Instagram
  • Privacy
  • Terms
  • Advertise
  • Contact Us
Monday, June 16
Facebook X (Twitter) Instagram YouTube
Technophile NewsTechnophile News
Demo
  • Home
  • News
  • PC
  • Phones
  • Android
  • Gadgets
  • Games
  • Guides
  • Accessories
  • Reviews
  • Spotlight
  • More
    • Artificial Intelligence
    • Web Stories
    • Press Release
Technophile NewsTechnophile News
Home » Databricks Has a Trick That Lets AI Models Improve Themselves
News

Databricks Has a Trick That Lets AI Models Improve Themselves

By News Room25 March 20253 Mins Read
Facebook Twitter Pinterest LinkedIn Telegram Tumblr Reddit WhatsApp Email
Share
Facebook Twitter LinkedIn Pinterest Email

Databricks, a company that helps big businesses build custom artificial intelligence models, has developed a machine learning trick that can boost the performance of an AI model without the need for clean labelled data.

Jonathan Frankle, chief AI scientist at Databricks, spent the past year talking to customers about the key challenges they face in getting AI to work reliably.

The problem, Frankle says, is dirty data.

”Everybody has some data, and has an idea of what they want to do,” Frankle says. But the lack of clean data makes it challenging to fine-tune a model to perform a specific task.. “Nobody shows up with nice, clean fine-tuning data that you can stick into a prompt or an [application programming interface],” for a model.

Databricks’ model could allow companies to eventually deploy their own agents to perform tasks, without data quality standing in the way.

The technique offers a rare look at some of the key tricks that engineers are now using to improve the abilities of advanced AI models, especially when good data is hard to come by. The method leverages ideas that have helped produce advanced reasoning models by combining reinforcement learning, a way for AI models to improve through practice, with “synthetic,” or AI-generated training data.

The latest models from OpenAI, Google, and DeepSeek all rely heavily on reinforcement learning as well as synthetic training data. WIRED revealed that Nvidia plans to acquire Gretel, a company that specializes in synthetic data. “We’re all navigating this space,” Frankle says.

The Databricks method exploits the fact that, given enough tries, even a weak model can score well on a given task or benchmark. Researchers call this method of boosting a model’s performance “best-of-N”. Databricks trained a model to predict which best-of-N result human testers would prefer, based on examples. The Databricks reward model, or DBRM, can then be used to improve the performance of other models without the need for further labelled data.

DBRM is then used to select the best outputs from a given model. This creates synthetic training data for further fine-tuning the model so that it produces a better output first time. Databricks calls its new approach Test-time Adaptive Optimization or TAO. “This method we’re talking about uses some relatively lightweight reinforcement learning to basically bake the benefits of best-of-N into the model itself,” Frankle says.

He adds that the research done by Databricks shows that the TAO method improves as it is scaled up to larger, more capable models. Reinforcement learning and synthetic data are already widely used but combining them in order to improve language models is a relatively new and technically challenging technique.

Databricks is unusually open about how it develops AI because it wants to show customers that it has the skills needed to create powerful custom models for them. The company previously revealed to WIRED how it developed DBX, a cutting-edge open source large language model (LLM) from scratch.

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email

Related News

Minnesota Shooting Suspect Allegedly Used Data Broker Sites to Find Targets’ Addresses

16 June 2025

The Definitive Story of Tesla Takedown

16 June 2025

Try This Free Version of Microsoft Office That Runs in Your Browser

16 June 2025

‘Psyop’: How Far-Right Conspiracy Theories About the Minnesota Shooting Evolved to Protect MAGA

16 June 2025

Companies Warn SEC That Mass Deportations Pose Serious Business Risk

16 June 2025

Justin Sun takes crypto company public — reportedly with help from the Trump Family

16 June 2025
Top Articles

Oppo Reno 14, Reno 14 Pro India Launch Timeline and Colourways Leaked

27 May 202542 Views

Vivo S30, Vivo S30 Pro Mini Launched With 6,500mAh Battery, 50-Megapixel Selfie Camera: Price, Specifications

29 May 202533 Views

How to Buy Ethical and Eco-Friendly Electronics

22 April 202532 Views
Stay In Touch
  • Facebook
  • YouTube
  • TikTok
  • WhatsApp
  • Twitter
  • Instagram
Don't Miss

Companies Warn SEC That Mass Deportations Pose Serious Business Risk

16 June 2025

Other filings suggested a recession could come even earlier. The community bank Hanmi Bank, under…

Justin Sun takes crypto company public — reportedly with help from the Trump Family

16 June 2025

Apple Risks Fresh EU Charge Sheet Over App Store Curbs

16 June 2025

How Apple Created a Custom iPhone Camera for F1

16 June 2025
Technophile News
Facebook X (Twitter) Instagram Pinterest YouTube Dribbble
  • Privacy Policy
  • Terms of use
  • Advertise
  • Contact Us
© 2025 Technophile News. All Rights Reserved.

Type above and press Enter to search. Press Esc to cancel.