Apple Shares Massive Dataset to Help Researchers Build Nano Banana-Like AI Models

Apple wants to help researchers and developers build better AI models with its dataset, even when it struggles to do so itself.

Advertisement
Written by Akash Dutta, Edited by Ketan Pratap | Updated: 29 October 2025 15:20 IST
Highlights
  • Apple’s Pico-Banana-400K is a dataset for text-guided image editing
  • The company used Nano Banana’s output to create the dataset
  • Apple’s dataset comes with a non-commercial research license

Apple said the dataset was created due to the absence of large-scale and openly accessible images

Photo Credit: Reuters

Apple researchers have released a large-scale dataset to help others develop image editing artificial intelligence (AI) models. Dubbed Pico-Banana-400K, the dataset contains 4,00,000 real images and their AI-edited counterparts that can be used to train large language models how to handle text-based image editing requests. It is an open-source dataset available with a research-only license, meaning it cannot be used for commercial purposes. Interestingly, the Cupertino-based tech giant's new dataset release comes at a time when it is struggling with native AI models itself.

Apple's Pico-Banana-400K Will Help Others Build Image Editing Models

A research paper titled “Pico-Banana-400K: A Large-Scale Dataset for Text-Guided Image Editing” was published on arXiv, an online journal. The dataset contains roughly 4,00,00 real photo edit pairs, built from OpenImages, organised into a 35-type edit taxonomy and split into single-turn edits, multi-turn sequences and preference pairs.

These design choices matter because they shift the training signal from synthetic, narrowly curated examples to instruction-rich, real-world scenarios that resemble what users actually ask for.

Advertisement

Pico-Banana-400K was produced by chaining a powerful generative model (Nano Banana) to create edits and another large multimodal model to act as an automated judge, filtering and retrying failed attempts. The result is a dataset emphasising photographic diversity, human-centric scenes and text-heavy shots. The photos also focus on nuance, with long and short instruction pairs to support research work.

Advertisement

Additionally, it also includes negative examples and preference pairs, which are crucial for alignment research and for teaching models not just what to do but what “better” looks like. The paper explicitly documents which edit types are robust (style transfers, global photometric changes) and which remain brittle (precise spatial relocations, text replacement on signs), making it unusually candid about limitations.

The dataset is currently available on GitHub, and can be used for any non-commercial use cases.

Advertisement

Interestingly, Apple has seemingly stalled with the company's in-house AI progress. While it has integrated the Apple Intelligence in more apps and features with the iPhone 17 series launch, the company continues to delay the Siri overhaul which was first announced in 2024.

 

For the latest tech news and reviews, follow Gadgets 360 on X, Facebook, WhatsApp, Threads and Google News. For the latest videos on gadgets and tech, subscribe to our YouTube channel. If you want to know everything about top influencers, follow our in-house Who'sThat360 on Instagram and YouTube.

Further reading: Apple, AI, Artificial Intelligence
Advertisement

Related Stories

Popular Mobile Brands
  1. Amazon Fire TV Stick 4K Select Launched in India With Vega OS
  2. Oppo Find X9 Series Confirmed to Be Available in India via Flipkart
  3. Vivo X300 Series Price, Key Features Leaked Ahead of Global Launch
  4. Nothing Phone 3a Lite Launched With Glyph Light At This Price
  5. TRAI, DoT Approve Presentation of Caller Names During Incoming Calls
  6. Nothing Phone 3a Lite First Impressions
  1. NASA’s X-59 Supersonic Jet Takes Historic First Flight, Paving Way for Quiet Supersonic Travel
  2. ASIC Clarifies Crypto Rules; Stablecoins, Tokenised Assets Flagged as Financial Products
  3. SpaceX Launches 28 Starlink Satellites, Lands Falcon 9 Booster in Pacific
  4. Idli Kadai, Starring Dhanush, Now Streaming on Netflix: What You Need to Know
  5. Ideabaaz Now Streaming on ZEE5: Everything You Need to Know
  6. Grey’s Anatomy Season 22 OTT Release: Know Where to Watch it Online?
  7. Bad Girl OTT Release Date: When and Where to Watch Tamil Drama Online?
  8. Adobe Partners With Google Cloud to Integrate Frontier AI Models Across Its Platforms
  9. Vivo X300, Vivo X300 Pro Price and Key Specifications Leaked Ahead of Global Launch
  10. OnePlus 15 India Launch Date Announced; to Debut as First Snapdragon 8 Elite Gen 5 Phone in India
Gadgets 360 is available in
Download Our Apps
Available in Hindi
© Copyright Red Pixels Ventures Limited 2025. All rights reserved.