Close Menu
    Facebook LinkedIn YouTube WhatsApp X (Twitter) Pinterest
    Trending
    • Tech Up Your Sourdough With These Upper-Crust Baking Gadgets
    • Resident Evil Requiem Revealed, but Where’s Leon Kennedy?
    • What Happens When You Remove the Filters from AI Love Generators?
    • First passenger flight for electric CTOL aircraft lands in JFK
    • Best Backpacking Tents (2025), WIRED-Tested and Reviewed
    • Microcurrent Devices: Do They Work and Are They Worth It? We Asked Skin Experts
    • Will Musk’s explosive row with Trump help or harm his businesses?
    • 7 AI Hentai Girlfriend Chat Websites No Filter
    Facebook LinkedIn WhatsApp
    Times FeaturedTimes Featured
    Saturday, June 7
    • Home
    • Founders
    • Startups
    • Technology
    • Profiles
    • Entrepreneurs
    • Leaders
    • Students
    • VC Funds
    • More
      • AI
      • Robotics
      • Industries
      • Global
    Times FeaturedTimes Featured
    Home»AI Technology News»DeepSeek might not be such good news for energy after all
    AI Technology News

    DeepSeek might not be such good news for energy after all

    Editor Times FeaturedBy Editor Times FeaturedJanuary 31, 2025No Comments3 Mins Read
    Facebook Twitter Pinterest Telegram LinkedIn Tumblr WhatsApp Email
    Share
    Facebook Twitter LinkedIn Pinterest Telegram Email WhatsApp Copy Link


    Add the truth that different tech companies, impressed by DeepSeek’s strategy, could now begin constructing their very own related low-cost reasoning fashions, and the outlook for power consumption is already looking so much much less rosy.

    The life cycle of any AI mannequin has two phases: coaching and inference. Coaching is the usually months-long course of through which the mannequin learns from information. The mannequin is then prepared for inference, which occurs every time anybody on the planet asks it one thing. Each often happen in information facilities, the place they require a lot of power to run chips and funky servers. 

    On the coaching facet for its R1 mannequin, DeepSeek’s workforce improved what’s known as a “combination of specialists” method, through which solely a portion of a mannequin’s billions of parameters—the “knobs” a mannequin makes use of to kind higher solutions—are turned on at a given time throughout coaching. Extra notably, they improved reinforcement studying, the place a mannequin’s outputs are scored after which used to make it higher. That is usually finished by human annotators, however the DeepSeek workforce received good at automating it. 

    The introduction of a approach to make coaching extra environment friendly may recommend that AI firms will use much less power to deliver their AI fashions to a sure commonplace. That’s probably not the way it works, although. 

    “⁠As a result of the worth of getting a extra clever system is so excessive,” wrote Anthropic cofounder Dario Amodei on his weblog, it “causes firms to spend extra, not much less, on coaching fashions.” If firms get extra for his or her cash, they may discover it worthwhile to spend extra, and due to this fact use extra power. “The positive aspects in price effectivity find yourself completely dedicated to coaching smarter fashions, restricted solely by the corporate’s monetary sources,” he wrote. It’s an instance of what’s referred to as the Jevons paradox.

    However that’s been true on the coaching facet so long as the AI race has been going. The power required for inference is the place issues get extra attention-grabbing. 

    DeepSeek is designed as a reasoning mannequin, which implies it’s meant to carry out properly on issues like logic, pattern-finding, math, and different duties that typical generative AI fashions wrestle with. Reasoning fashions do that utilizing one thing known as “chain of thought.” It permits the AI mannequin to interrupt its job into components and work by them in a logical order earlier than coming to its conclusion. 

    You’ll be able to see this with DeepSeek. Ask whether or not it’s okay to lie to guard somebody’s emotions, and the mannequin first tackles the query with utilitarianism, weighing the quick good towards the potential future hurt. It then considers Kantian ethics, which suggest that you must act based on maxims that might be common legal guidelines. It considers these and different nuances earlier than sharing its conclusion. (It finds that mendacity is “typically acceptable in conditions the place kindness and prevention of hurt are paramount, but nuanced with no common answer,” for those who’re curious.)



    Source link

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Editor Times Featured
    • Website

    Related Posts

    Manus has kick-started an AI agent boom in China

    June 5, 2025

    What’s next for AI and math

    June 4, 2025

    Inside the tedious effort to tally AI’s energy appetite

    June 3, 2025

    Fueling seamless AI at scale

    May 30, 2025

    This benchmark used Reddit’s AITA to test how much AI models suck up to us

    May 30, 2025

    Designing Pareto-optimal GenAI workflows with syftr

    May 28, 2025

    Comments are closed.

    Editors Picks

    Tech Up Your Sourdough With These Upper-Crust Baking Gadgets

    June 7, 2025

    Resident Evil Requiem Revealed, but Where’s Leon Kennedy?

    June 7, 2025

    What Happens When You Remove the Filters from AI Love Generators?

    June 7, 2025

    First passenger flight for electric CTOL aircraft lands in JFK

    June 7, 2025
    Categories
    • Founders
    • Startups
    • Technology
    • Profiles
    • Entrepreneurs
    • Leaders
    • Students
    • VC Funds
    About Us
    About Us

    Welcome to Times Featured, an AI-driven entrepreneurship growth engine that is transforming the future of work, bridging the digital divide and encouraging younger community inclusion in the 4th Industrial Revolution, and nurturing new market leaders.

    Empowering the growth of profiles, leaders, entrepreneurs businesses, and startups on international landscape.

    Asia-Middle East-Europe-North America-Australia-Africa

    Facebook LinkedIn WhatsApp
    Featured Picks

    How ‘Based’ Is Grok 3? + Robinhood C.E.O. Vlad Tenev on Markets for Everything + Vibecoding 101

    February 21, 2025

    11 Methods and Hardware Tools for 3D Scanning

    January 31, 2025

    Unusual cantilevered tower thinks outside the box

    June 4, 2025
    Categories
    • Founders
    • Startups
    • Technology
    • Profiles
    • Entrepreneurs
    • Leaders
    • Students
    • VC Funds
    Copyright © 2024 Timesfeatured.com IP Limited. All Rights.
    • Privacy Policy
    • Disclaimer
    • Terms and Conditions
    • About us
    • Contact us

    Type above and press Enter to search. Press Esc to cancel.