Close Menu
  • Home
  • AI Models
    • DeepSeek
    • xAI
    • OpenAI
    • Meta AI Llama
    • Google DeepMind
    • Amazon AWS AI
    • Microsoft AI
    • Anthropic (Claude)
    • NVIDIA AI
    • IBM WatsonX Granite 3.1
    • Adobe Sensi
    • Hugging Face
    • Alibaba Cloud (Qwen)
    • Baidu (ERNIE)
    • C3 AI
    • DataRobot
    • Mistral AI
    • Moonshot AI (Kimi)
    • Google Gemma
    • xAI
    • Stability AI
    • H20.ai
  • AI Research
    • Allen Institue for AI
    • arXiv AI
    • Berkeley AI Research
    • CMU AI
    • Google Research
    • Microsoft Research
    • Meta AI Research
    • OpenAI Research
    • Stanford HAI
    • MIT CSAIL
    • Harvard AI
  • AI Funding & Startups
    • AI Funding Database
    • CBInsights AI
    • Crunchbase AI
    • Data Robot Blog
    • TechCrunch AI
    • VentureBeat AI
    • The Information AI
    • Sifted AI
    • WIRED AI
    • Fortune AI
    • PitchBook
    • TechRepublic
    • SiliconANGLE – Big Data
    • MIT News
    • Data Robot Blog
  • Expert Insights & Videos
    • Google DeepMind
    • Lex Fridman
    • Matt Wolfe AI
    • Yannic Kilcher
    • Two Minute Papers
    • AI Explained
    • TheAIEdge
    • Matt Wolfe AI
    • The TechLead
    • Andrew Ng
    • OpenAI
  • Expert Blogs
    • François Chollet
    • Gary Marcus
    • IBM
    • Jack Clark
    • Jeremy Howard
    • Melanie Mitchell
    • Andrew Ng
    • Andrej Karpathy
    • Sebastian Ruder
    • Rachel Thomas
    • IBM
  • AI Policy & Ethics
    • ACLU AI
    • AI Now Institute
    • Center for AI Safety
    • EFF AI
    • European Commission AI
    • Partnership on AI
    • Stanford HAI Policy
    • Mozilla Foundation AI
    • Future of Life Institute
    • Center for AI Safety
    • World Economic Forum AI
  • AI Tools & Product Releases
    • AI Assistants
    • AI for Recruitment
    • AI Search
    • Coding Assistants
    • Customer Service AI
    • Image Generation
    • Video Generation
    • Writing Tools
    • AI for Recruitment
    • Voice/Audio Generation
  • Industry Applications
    • Finance AI
    • Healthcare AI
    • Legal AI
    • Manufacturing AI
    • Media & Entertainment
    • Transportation AI
    • Education AI
    • Retail AI
    • Agriculture AI
    • Energy AI
  • AI Art & Entertainment
    • AI Art News Blog
    • Artvy Blog » AI Art Blog
    • Weird Wonderful AI Art Blog
    • The Chainsaw » AI Art
    • Artvy Blog » AI Art Blog
What's Hot

AI makes us impotent

Stanford HAI’s 2025 AI Index Reveals Record Growth in AI Capabilities, Investment, and Regulation

New MIT CSAIL study suggests that AI won’t steal as many jobs as expected

Facebook X (Twitter) Instagram
Advanced AI News
  • Home
  • AI Models
    • Adobe Sensi
    • Aleph Alpha
    • Alibaba Cloud (Qwen)
    • Amazon AWS AI
    • Anthropic (Claude)
    • Apple Core ML
    • Baidu (ERNIE)
    • ByteDance Doubao
    • C3 AI
    • Cohere
    • DataRobot
    • DeepSeek
  • AI Research & Breakthroughs
    • Allen Institue for AI
    • arXiv AI
    • Berkeley AI Research
    • CMU AI
    • Google Research
    • Meta AI Research
    • Microsoft Research
    • OpenAI Research
    • Stanford HAI
    • MIT CSAIL
    • Harvard AI
  • AI Funding & Startups
    • AI Funding Database
    • CBInsights AI
    • Crunchbase AI
    • Data Robot Blog
    • TechCrunch AI
    • VentureBeat AI
    • The Information AI
    • Sifted AI
    • WIRED AI
    • Fortune AI
    • PitchBook
    • TechRepublic
    • SiliconANGLE – Big Data
    • MIT News
    • Data Robot Blog
  • Expert Insights & Videos
    • Google DeepMind
    • Lex Fridman
    • Meta AI Llama
    • Yannic Kilcher
    • Two Minute Papers
    • AI Explained
    • TheAIEdge
    • Matt Wolfe AI
    • The TechLead
    • Andrew Ng
    • OpenAI
  • Expert Blogs
    • François Chollet
    • Gary Marcus
    • IBM
    • Jack Clark
    • Jeremy Howard
    • Melanie Mitchell
    • Andrew Ng
    • Andrej Karpathy
    • Sebastian Ruder
    • Rachel Thomas
    • IBM
  • AI Policy & Ethics
    • ACLU AI
    • AI Now Institute
    • Center for AI Safety
    • EFF AI
    • European Commission AI
    • Partnership on AI
    • Stanford HAI Policy
    • Mozilla Foundation AI
    • Future of Life Institute
    • Center for AI Safety
    • World Economic Forum AI
  • AI Tools & Product Releases
    • AI Assistants
    • AI for Recruitment
    • AI Search
    • Coding Assistants
    • Customer Service AI
    • Image Generation
    • Video Generation
    • Writing Tools
    • AI for Recruitment
    • Voice/Audio Generation
  • Industry Applications
    • Education AI
    • Energy AI
    • Finance AI
    • Healthcare AI
    • Legal AI
    • Media & Entertainment
    • Transportation AI
    • Manufacturing AI
    • Retail AI
    • Agriculture AI
  • AI Art & Entertainment
    • AI Art News Blog
    • Artvy Blog » AI Art Blog
    • Weird Wonderful AI Art Blog
    • The Chainsaw » AI Art
    • Artvy Blog » AI Art Blog
Advanced AI News
Home » Exclusive: Panzura unlocks metadata from IBM Deep Archive files for AI training
IBM

Exclusive: Panzura unlocks metadata from IBM Deep Archive files for AI training

Advanced AI BotBy Advanced AI BotApril 2, 2025No Comments4 Mins Read
Share Facebook Twitter Pinterest Copy Link Telegram LinkedIn Tumblr Email
Share
Facebook Twitter LinkedIn Pinterest Email


Panzura LLC, maker of a hybrid cloud data management platform, announced today that its Symphony data services platform is now integrated with IBM Corp.’s Storage Deep Archive.

The integration, on the IBM Diamondback tape library (pictured), makes “cold” data — meaning archived to tape — easily accessible for information discovery and artificial intelligence training. Symphony is a cloud-based single management point for distributed file systems that span multiple locations. It automates file and object data movement, data discovery and assessment and structures unstructured data using metadata. The company specializes in helping large organizations manage, store and access unstructured data like files, images and videos.

IBM Storage Deep Archive is a cloud-based, low-cost storage tier for long-term data retention and digital preservation. It features air-gapped and encrypted storage to protect against cyberthreats. Deep Archive makes data stored on tape accessible without specialized technical expertise.

Symphony’s data movement framework is also integrated with IBM’s Fusion Data Catalog to create a massively scalable data fabric for rapidly ingesting metadata in a business context. Fusion Data Catalog acts as a centralized inventory of data assets with automatic discovery and classification, data lineage tracking and data governance policies.

Metadata discovery

Symphony’s metadata discovery features unlock information that would otherwise be inaccessible without human inspection and tagging, said Mike Harvey, senior vice president of technology at Panzura and founder of Moonwalk Universal Inc. Panzura acquired that company, which makes data assessment and storage optimization technology, last year.

“There are a lot of workflows in the enterprise where applications, knowledge workers and scientific users reacquire entire data artifacts when really what they’re interested in is metadata,” he said. “Without cataloging or a metadata services layer, they have to re-acquire entire artifacts to open an application.”

Automatic data discovery scans files for recognizable patterns such as phone numbers and can extract them as metadata. The software can also look for information that is likely to be adjacent to other data types, such as phone numbers and addresses.

“A chief legal counsel can ask a question like how many files the organization has that contain Social Security numbers,” said Glen Shok, vice president of strategic alliances at Panzura. “Those files could be in SharePoint, an S3 bucket, on a NetApp file share and on a network file system. We deliver them all through a chat window.”

Single namespace

Metadata used to train AI models is often embedded in file headers or is otherwise invisible to file systems, Shok said. Symphony extracts that metadata and makes files retrievable as if they were in a local file system.

“We can stream the content to [Amazon Web Services Inc.’s] Glacier Flexible Retrieval, which front-ends the tape infrastructure, without it looking like the data set has moved,” he said. “The namespace and the security is preserved at the front-end.”

In addition to low cost, Deep Archive’s high-density storage reduces energy consumption by up to 97% compared to hard disks, Harvey said.

“You get the carbon offset, low energy consumption, long-term retention and protection from ransomware without severing the link to the original namespace,” he said.

The combination is expected to be especially useful in training AI models, which require vast amounts of data. Information in archival storage often lacks rich metadata and is slow to retrieve. Panzura said integration with Deep Archive can make that data useful for purposes like retrieval-augmented generation.

“We’re creating a metadata catalog from over 500 different file types that’s relevant to the user base,” Shok said. “Instead of building large language models, you can build small language models for expert systems specific to engineering or accounting.”

Photo: IBM

Your vote of support is important to us and it helps us keep the content FREE.

One click below supports our mission to provide free, deep, and relevant content.  

Join our community on YouTube

Join the community that includes more than 15,000 #CubeAlumni experts, including Amazon.com CEO Andy Jassy, Dell Technologies founder and CEO Michael Dell, Intel CEO Pat Gelsinger, and many more luminaries and experts.

“TheCUBE is an important partner to the industry. You guys really are a part of our events and we really appreciate you coming and I know people appreciate the content you create as well” – Andy Jassy

THANK YOU



Source link

Follow on Google News Follow on Flipboard
Share. Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link
Previous ArticleAustralia’s Synchron uses brain signals to train foundation AI model — Capital Brief
Next Article Tony Blair Institute AI copyright report sparks backlash
Advanced AI Bot
  • Website

Related Posts

Will the Launch of watsonx AI Labs Be a Game Changer for IBM? – June 5, 2025

June 6, 2025

IBM Endicott’s Amazing Vanishing Act

June 6, 2025

IBM’s cloud crisis deepens: 54 services disrupted in latest outage

June 5, 2025
Leave A Reply Cancel Reply

Latest Posts

Men’s Swimwear Gets Casual At Miami Swim Week 2025

Original Prototype for Jane Birkin’s Hermes Bag Consigned to Sotheby’s

Viral Trump Vs. Musk Feud Ignites A Meme Chain Reaction

UK Art Dealer Sentenced To 2.5 Years In Jail For Selling Art to Suspected Hezbollah Financier

Latest Posts

AI makes us impotent

June 7, 2025

Stanford HAI’s 2025 AI Index Reveals Record Growth in AI Capabilities, Investment, and Regulation

June 7, 2025

New MIT CSAIL study suggests that AI won’t steal as many jobs as expected

June 7, 2025

Subscribe to News

Subscribe to our newsletter and never miss our latest news

Subscribe my Newsletter for New Posts & tips Let's stay updated!

Welcome to Advanced AI News—your ultimate destination for the latest advancements, insights, and breakthroughs in artificial intelligence.

At Advanced AI News, we are passionate about keeping you informed on the cutting edge of AI technology, from groundbreaking research to emerging startups, expert insights, and real-world applications. Our mission is to deliver high-quality, up-to-date, and insightful content that empowers AI enthusiasts, professionals, and businesses to stay ahead in this fast-evolving field.

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

Subscribe my Newsletter for New Posts & tips Let's stay updated!

YouTube LinkedIn
  • Home
  • About Us
  • Advertise With Us
  • Contact Us
  • DMCA
  • Privacy Policy
  • Terms & Conditions
© 2025 advancedainews. Designed by advancedainews.

Type above and press Enter to search. Press Esc to cancel.