Close Menu
    Facebook LinkedIn YouTube WhatsApp X (Twitter) Pinterest
    Trending
    • Portable water filter provides safe drinking water from any source
    • MAGA Is Increasingly Convinced the Trump Assassination Attempt Was Staged
    • NCAA seeks faster trial over DraftKings disputed March Madness branding case
    • AI Trusted Less Than Social Media and Airlines, With Grok Placing Last, Survey Says
    • Extragalactic Archaeology tells the ‘life story’ of a whole galaxy
    • Swedish semiconductor startup AlixLabs closes €15 million Series A to scale atomic-level etching technology
    • Republican Mutiny Sinks Trump’s Push to Extend Warrantless Surveillance
    • Yocha Dehe slams Vallejo Council over rushed casino deal approval process
    Facebook LinkedIn WhatsApp
    Times FeaturedTimes Featured
    Saturday, April 18
    • Home
    • Founders
    • Startups
    • Technology
    • Profiles
    • Entrepreneurs
    • Leaders
    • Students
    • VC Funds
    • More
      • AI
      • Robotics
      • Industries
      • Global
    Times FeaturedTimes Featured
    Home»AI Technology News»This is where the data to build AI comes from
    AI Technology News

    This is where the data to build AI comes from

    Editor Times FeaturedBy Editor Times FeaturedDecember 29, 2024No Comments2 Mins Read
    Facebook Twitter Pinterest Telegram LinkedIn Tumblr WhatsApp Email
    Share
    Facebook Twitter LinkedIn Pinterest Telegram Email WhatsApp Copy Link


    Their findings, shared exclusively with MIT Technology Review, present a worrying development: AI’s information practices threat concentrating energy overwhelmingly within the arms of some dominant know-how corporations. 

    Within the early 2010s, information units got here from a wide range of sources, says Shayne Longpre, a researcher at MIT who’s a part of the mission. 

    It got here not simply from encyclopedias and the net, but in addition from sources corresponding to parliamentary transcripts, incomes calls, and climate experiences. Again then, AI information units have been particularly curated and picked up from completely different sources to swimsuit particular person duties, Longpre says.

    Then transformers, the structure underpinning language fashions, have been invented in 2017, and the AI sector began seeing efficiency get higher the larger the fashions and information units have been. At the moment, most AI information units are constructed by indiscriminately hoovering materials from the web. Since 2018, the net has been the dominant supply for information units utilized in all media, corresponding to audio, photos, and video, and a spot between scraped information and extra curated information units has emerged and widened.

    “In basis mannequin improvement, nothing appears to matter extra for the capabilities than the dimensions and heterogeneity of the information and the net,” says Longpre. The necessity for scale has additionally boosted the usage of artificial information massively.

    The previous few years have additionally seen the rise of multimodal generative AI fashions, which may generate movies and pictures. Like giant language fashions, they want as a lot information as attainable, and one of the best supply for that has develop into YouTube. 

    For video fashions, as you’ll be able to see on this chart, over 70% of information for each speech and picture information units comes from one supply.

    This might be a boon for Alphabet, Google’s mother or father firm, which owns YouTube. Whereas textual content is distributed throughout the net and managed by many alternative web sites and platforms, video information is extraordinarily concentrated in a single platform.



    Source link

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Editor Times Featured
    • Website

    Related Posts

    How robots learn: A brief, contemporary history

    April 17, 2026

    Vibe Coding Best Practices: 5 Claude Code Habits

    April 16, 2026

    Why having “humans in the loop” in an AI war is an illusion

    April 16, 2026

    Making AI operational in constrained public sector environments

    April 16, 2026

    Treating enterprise AI as an operating layer

    April 16, 2026

    Building trust in the AI era with privacy-led UX

    April 15, 2026

    Comments are closed.

    Editors Picks

    Portable water filter provides safe drinking water from any source

    April 18, 2026

    MAGA Is Increasingly Convinced the Trump Assassination Attempt Was Staged

    April 18, 2026

    NCAA seeks faster trial over DraftKings disputed March Madness branding case

    April 18, 2026

    AI Trusted Less Than Social Media and Airlines, With Grok Placing Last, Survey Says

    April 18, 2026
    Categories
    • Founders
    • Startups
    • Technology
    • Profiles
    • Entrepreneurs
    • Leaders
    • Students
    • VC Funds
    About Us
    About Us

    Welcome to Times Featured, an AI-driven entrepreneurship growth engine that is transforming the future of work, bridging the digital divide and encouraging younger community inclusion in the 4th Industrial Revolution, and nurturing new market leaders.

    Empowering the growth of profiles, leaders, entrepreneurs businesses, and startups on international landscape.

    Asia-Middle East-Europe-North America-Australia-Africa

    Facebook LinkedIn WhatsApp
    Featured Picks

    Ex-Reuters team out of Denmark raises €1.5 million for AI newsroom Financial News System

    March 30, 2026

    M5 MacBook Air vs. M4, M3, M2, M1: Should You Upgrade?

    March 18, 2026

    How to Protest Safely in the Age of Surveillance

    June 12, 2025
    Categories
    • Founders
    • Startups
    • Technology
    • Profiles
    • Entrepreneurs
    • Leaders
    • Students
    • VC Funds
    Copyright © 2024 Timesfeatured.com IP Limited. All Rights.
    • Privacy Policy
    • Disclaimer
    • Terms and Conditions
    • About us
    • Contact us

    Type above and press Enter to search. Press Esc to cancel.