Close Menu

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    DOJ did not classify MNGO as a commodity

    September 11, 2026

    Michael Kosta Nails Right-Wing Mamdani 9/11 Meltdown With 1 Key Observation

    September 11, 2026

    Two Truths and a Lie About Digital PR

    September 11, 2026
    Facebook X (Twitter) Instagram
    Trending
    • DOJ did not classify MNGO as a commodity
    • Michael Kosta Nails Right-Wing Mamdani 9/11 Meltdown With 1 Key Observation
    • Two Truths and a Lie About Digital PR
    • Online Reputation Management Services – Company Profile
    • A Short History of Trend-Following and Momentum
    • Nicolle Wallace Damns Trump With A Story About George W. Bush And A Spoon
    • Common AC Mistakes That Could Be Costing You Money
    • More Than 100,000 Demand FIFA Investigate Its President’s Ties To Trump
    Facebook X (Twitter)
    SBM Global News
    Demo
    • Home
    • Top Stories
      • Politics
    • Business
      • Small Business
      • Marketing
    • Finance
      • Investment
    • Technology

      Online Reputation Management Services – Company Profile

      September 11, 2026
      Read More

      AI research startup Listen Labs scrubbed a $1.5B funding round for Salesforce talks

      September 10, 2026
      Read More

      Google Cloud races to catch up in the AI deployment wars with Accenture deal

      September 9, 2026
      Read More

      Phil Schiller’s App Store exit reportedly driven by wariness over future plans

      September 7, 2026
      Read More

      Clucky’s new alarm app wakes you up with a crowing rooster

      September 6, 2026
      Read More
    • Lifestyle
      • Travel
    • Feel Good
    • Get In Touch
    SBM Global News
    Demo
    Home»Technology»Google Deepmind trains a video game-playing AI to be your co-op companion
    Technology

    Google Deepmind trains a video game-playing AI to be your co-op companion

    By Staff WriterMarch 14, 20245 Mins Read
    Facebook Twitter LinkedIn Reddit Email
    #image_title
    Share
    Facebook Twitter LinkedIn Pinterest Email

    AI models that play games go back decades, but they generally specialize in one game and always play to win. Google Deepmind researchers have a different goal with their latest creation: a model that learned to play multiple 3D games like a human, but also does its best to understand and act on your verbal instructions.

    There are of course “AI” or computer characters that can do this kind of thing, but they’re more like features of a game: NPCs that you can use formal in-game commands to indirectly control.

    Deepmind’s SIMA (scalable instructable multiworld agent) doesn’t have any kind of access to the game’s internal code or rules; instead, it was trained on many, many hours of video showing gameplay by humans. From this data — and the annotations provided by data labelers — the model learns to associate certain visual representations of actions, objects, and interactions. They also recorded videos of players instructing one another to do things in game.

    For example, it might learn from how the pixels move in a certain pattern on screen that this is an action called “moving forward,” or when the character approaches a door-like object and uses the doorknob-looking object, that’s “opening” a “door.” Simple things like that, tasks or events that take a few seconds but are more than just pressing a key or identifying something.

    The training videos were taken in multiple games, from Valheim to Goat Simulator 3, the developers of which were involved with and consenting to this use of their software. One of the main goals, the researchers said in a call with press, was to see whether training an AI to play one set of games makes it capable of playing others it hasn’t seen, a process called generalization.

    The answer is yes, with caveats. AI agents trained on multiple games performed better on games they hadn’t been exposed to. But of course many games involve specific and unique mechanics or terms that will stymie the best-prepared AI. But there’s nothing stopping the model from learning those except a lack of training data.

    This is partly because, although there is lots of in-game lingo, there really are only so many “verbs” players have that really affect the game world. Whether you’re assembling a lean-to, pitching a tent, or summoning a magical shelter, you’re really “building a house,” right? So this map of several dozen primitives the agent currently recognizes is really interesting to peruse:

    A map of several dozen actions SIMA recognizes and can perform or combine.

    The researchers’ ambition, on top of advancing the ball in agent-based AI fundamentally, is to create a more natural game-playing companion than the stiff, hard-coded ones we have today.

    “Rather than having a superhuman agent you play against, you can have SIMA players beside you that are cooperative, that you can give instructions to,” said Tim Harley, one of the proejct’s leads.

    Since when they’re playing, all they see is the pixels of the game screen, they have to learn how to do stuff in much the same way we do — but it also means they can adapt and produce emergent behaviors as well.

    You may be curious how this stacks up against a common method of making agent-type AIs, the simulator approach, in which a mostly unsupervised model experiments wildly in a 3D simulated world running far faster than real time, allowing it to learn the rules intuitively and design behaviors around them without nearly as much annotation work.

    “Traditional simulator-based agent training uses reinforcement learning for training, which requires the game or environment to provide a ‘reward’ signal for the agent to learn from – for example win/loss in the case of Go or Starcraft, or ‘score’ for Atari,” Harley told TechCrunch, and noted that this approach was used for those games and produced phenomenal results.

    Demo

    “In the games that we use, such as the commercial games from our partners,” he continued, “We do not have access to such a reward signal. Moreover, we are interested in agents that can do a wide variety of tasks described in open-ended text – it’s not feasible for each game to evaluate a ‘reward’ signal for each possible goal. Instead, we train agents using imitation learning from human behavior, given goals in text.”

    In other words, having a strict reward structure can limit the agent in what it pursues, since if it is guided by score it will never attempt anything that does not maximize that value. But if it values something more abstract, like how close its action is to one it has observed working before, it can be trained to “want” to do almost anything as long as the training data represents it somehow.

    Other companies are looking into this kind of open-ended collaboration and creation as well; conversations with NPCs are being looked at pretty hard as opportunities to put an LLM-type chatbot to work, for instance. And simple improvised actions or interactions are also being simulated and tracked by AI in some really interesting research into agents.

    Of course there are also the experiments into infinite games like MarioGPT, but that’s another matter entirely.

    View original article here

    Share. Facebook Twitter LinkedIn Email Reddit
    Previous ArticleA Ban? A Sale? The Big Questions Hanging Over TikTok
    Next Article 20 Email Opt-In Examples I Love (For Your Inspiration)

    Related Posts

    Online Reputation Management Services – Company Profile

    September 11, 2026
    Read More

    AI research startup Listen Labs scrubbed a $1.5B funding round for Salesforce talks

    September 10, 2026
    Read More

    Google Cloud races to catch up in the AI deployment wars with Accenture deal

    September 9, 2026
    Read More
    Add A Comment

    Leave A Reply Cancel Reply

    Demo
    Top Posts

    Former FBI, CIA Head Has ‘Serious Concerns’ With Trump Cabinet Picks

    December 28, 2024435

    Emirates to operate next-gen A350 on the third daily service to Cape Town

    January 14, 2026256

    AAVE Price Prediction: Target $215-225 by Mid-January 2025 as Technical Indicators Signal Bullish Momentum

    December 15, 2025240

    Wall Street Slumps As Alphabet, Qualcomm And Other Tech Stocks Fall

    February 7, 2026212
    Don't Miss
    Investment

    DOJ did not classify MNGO as a commodity

    By Staff WriterSeptember 11, 20263 Mins Read

    Avraham Eisenberg was arrested in Puerto Rico on Dec. 26 on commodities fraud and manipulation…

    Read More

    Michael Kosta Nails Right-Wing Mamdani 9/11 Meltdown With 1 Key Observation

    September 11, 2026

    Two Truths and a Lie About Digital PR

    September 11, 2026

    Online Reputation Management Services – Company Profile

    September 11, 2026
    Stay In Touch
    • Facebook
    • Twitter
    Demo
    About Us

    Small Business Minder brings together business and related news from around the world in one place. Follow us for all the business news you'll need.

    Facebook X (Twitter)
    Our Picks

    DOJ did not classify MNGO as a commodity

    September 11, 2026

    Michael Kosta Nails Right-Wing Mamdani 9/11 Meltdown With 1 Key Observation

    September 11, 2026
    Most Popular

    Former FBI, CIA Head Has ‘Serious Concerns’ With Trump Cabinet Picks

    December 28, 2024435

    Emirates to operate next-gen A350 on the third daily service to Cape Town

    January 14, 2026256
    © 2026 Small Business Minder
    • Home
    • Get In Touch

    Type above and press Enter to search. Press Esc to cancel.

    Ad Blocker Enabled!
    Ad Blocker Enabled!
    Our website is made possible by displaying online advertisements to our visitors. To get the most from our site, please disable your Ad Blocker.