Close Menu
    Facebook LinkedIn YouTube WhatsApp X (Twitter) Pinterest
    Trending
    • These Were My Favorite Things Samsung Unpacked During Its 2026 Galaxy Event
    • AI minister role boosted but tech department axed in Burnham shake-up
    • Loop Engineering for RAG Question Parsing: The Small Loop That Runs Before Retrieval
    • The risk of weather data sabotage is rising
    • Hand-E Now Reaches 100 mm Without Giving Up an Ounce of Precision
    • Weight loss drug effectiveness and long term maintenance
    • Here’s what Albo’s ‘Office of AI’ means for Australian tech
    • YouTube and X Have Become ‘Gateways’ to Nudify Apps
    Facebook LinkedIn WhatsApp
    Times FeaturedTimes Featured
    Thursday, July 23
    • Home
    • Founders
    • Startups
    • Technology
    • Profiles
    • Entrepreneurs
    • Leaders
    • Students
    • VC Funds
    • More
      • AI
      • Robotics
      • Industries
      • Global
    Times FeaturedTimes Featured
    Home»Artificial Intelligence»Tree of Thought Prompting: Teaching LLMs to Think Slowly
    Artificial Intelligence

    Tree of Thought Prompting: Teaching LLMs to Think Slowly

    Editor Times FeaturedBy Editor Times FeaturedMay 29, 2025No Comments8 Mins Read
    Facebook Twitter Pinterest Telegram LinkedIn Tumblr WhatsApp Email
    Share
    Facebook Twitter LinkedIn Pinterest Telegram Email WhatsApp Copy Link


    “Sluggish pondering (system 2) is effortful, rare, logical, calculating, aware.”

    — Daniel Kahneman, Pondering, Quick and Sluggish

    Giant Language fashions (LLMs) similar to ChatGPT usually behave as what Daniel Kahneman—winner of the 2002 Nobel Memorial Prize in Financial Sciences—defines as System 1: it’s quick, assured and virtually easy. You might be proper: there’s additionally a System 2, a slower, extra effortful mode of pondering.

    Researchers have been creating strategies these years to carry System 2-style pondering into LLMs by means of higher prompting methods. Tree-of-Thought (ToT) prompting is one probably the most excellent strategies and it permits the mannequin to comply with a number of reasoning paths resulting in probably higher choices.

    On this weblog put up, I’ll do a case examine the place a ToT-powered LLM agent performs the traditional recreation Minesweeper — not by guessing, however by reasoning. Similar to you’ll.

    CoT and ToT

    We’d begin our story with Chain-of-Thought (CoT) prompting — a way that guides LLMs to motive step-by-step. Have a look at the next instance:

    Q: I purchased 10 apples within the grocery store. Then I gave 2 apples to the neighbor and one other 2 to my pigs. Not sufficient apples for me now!!I then purchased 5 extra apples from a grocery however 1 was rotten. What number of apples now?
    A: You used to have 10 apples. You gave 2 apples to the neighbor and a pair of to the pigs. Now you may have 6 apples. Then you definitely purchased 5 extra apples, so that you had 11 apples. Lastly, you ate 1 apple, so you continue to have 10 apples in the long run.

    As you’ll be able to see, guided to suppose step-by-step, the LLM simulates higher reasoning.

    Tree-of-Thought (ToT) prompting expands on CoT. Because the identify suggests, It organizes reasoning like a tree the place every node is a possible “thought,” and branches are potential paths. It doesn’t comply with a linear course of like CoT.

    In observe, ToT creates a tree construction with branches, every with sub-steps resulting in a ultimate decision. The mannequin then evaluates every step, assigning every thought a classification as “positive”, “probably” or “unattainable”. Subsequent, ToT goes by means of your complete drawback house by making use of search algorithms similar to Breadth-First Search (BFS) and Depth-First Search (DFS) in an effort to select one of the best paths.

    Right here is an instance in an workplace constructing: I’ve an workplace with a max capability of 20 individuals, however 28 persons are coming this week. I’ve the next three branches as potential options:

    • Department 1: Transfer 8 individuals
      • Is there one other room close by? → Sure
      • Does it have house for 8 extra individuals? → Sure
      • Can we transfer individuals with none administrative course of? → Perhaps
      • Analysis: Promising!
    • Department 2: Increase the Room 
      • Can we make the room bigger? → Perhaps
      • Is that this allowed below security? → No
      • Can we request an exception within the constructing? → No
      • Analysis: It wont work!
    • Department 3: Break up the group into two
      • Can we divide these individuals into two teams? → Sure
      • Can we allow them to come on completely different days? → Perhaps
      • Analysis: Good potential!

    As you’ll be able to see, this course of mimics how we clear up laborious issues: we don’t suppose in a straight line. As an alternative, we discover, consider, and select.

    Case examine

    Minesweeper

    Minesweeper: wikipedia 

    It’s virtually unattainable however In case you don’t know, minesweeper is an easy online game.

    The board is split into cells, with mines randomly distributed. The quantity on a cell reveals the variety of mines adjoining to it. You win whenever you open all of the cells. Nonetheless, for those who hit a mine earlier than opening all of the cells, the sport is over and also you lose.

    We’re making use of ToT in Minesweeper which requires some logic guidelines and reasoning below constraints.

    We simulate the sport with the next code:

    # --- Sport Setup ---
    def generate_board():
       board = np.zeros((BOARD_SIZE, BOARD_SIZE), dtype=int)
       mines = set()
       whereas len(mines) < NUM_MINES:
           r, c = random.randint(0, BOARD_SIZE-1), random.randint(0, BOARD_SIZE-1)
           if (r, c) not in mines:
               mines.add((r, c))
               board[r][c] = -1  # -1 represents a mine
    
    
       # Fill in adjoining mine counts
       for r in vary(BOARD_SIZE):
           for c in vary(BOARD_SIZE):
               if board[r][c] == -1:
                   proceed
               depend = 0
               for dr in [-1, 0, 1]:
                   for dc in [-1, 0, 1]:
                       if 0 <= r+dr < BOARD_SIZE and 0 <= c+dc < BOARD_SIZE:
                           if board[r+dr][c+dc] == -1:
                               depend += 1
               board[r][c] = depend
       return board, mines

    You possibly can see that we generate a BOARD_SIZE*BOARD_SIZE measurement board with NUM_MINES mines.

    ToT LLM Agent

    We at the moment are able to construct our ToT LLM agent to resolve the puzzle of minesweeper. First, we have to outline a operate that returns thought on the present board by utilizing an LLM similar to GPT-4o.

    def llm_generate_thoughts(board, revealed, flagged_mines, known_safe, ok=3):
    
       board_text = board_to_text(board, revealed)
    
       valid_moves = [[r, c] for r in vary(BOARD_SIZE) for c in vary(BOARD_SIZE) if not revealed[r][c] and [r, c] not in flagged_mines]
    
       immediate = f"""
    
    You might be taking part in a 8x8 Minesweeper recreation.
    
    - A quantity (0–10) reveals what number of adjoining mines a revealed cell has.
    
    - A '?' means the cell is hidden.
    
    - You might have flagged these mines: {flagged_mines}
    
    - You realize these cells are secure: {known_safe}
    
    - Your job is to decide on ONE hidden cell that's least more likely to include a mine.
    
    - Use the next logic:
    
     - If a cell reveals '1' and touches precisely one '?', that cell should be a mine.
    
     - If a cell reveals '1' and touches one already flagged mine, different neighbors are secure.
    
     - Cells subsequent to '0's are typically secure.
    
    You might have the next board:
    
    {board_text}
    
    Listed here are all legitimate hidden cells you'll be able to select from:
    
    {valid_moves}
    
    Step-by-step:
    
    1. Listing {ok} potential cells to click on subsequent.
    
    2. For every, clarify why it may be secure (primarily based on adjoining numbers and identified information).
    
    3. Charge every transfer from 0.0 to 1.0 as a security rating (1 = positively secure).
    
    Return your reply on this actual JSON format:
    
    [
    
     {{ "cell": [row, col], "motive": "...", "rating": 0.95 }},
    
     ...
    
    ]
    
    """
    
       attempt:
    
           response = shopper.chat.completions.create(
    
               mannequin="gpt-4o",
    
               messages=[{"role": "user", "content": prompt}],
    
               temperature=0.3,
    
           )
    
           content material = response.selections[0].message.content material.strip()
    
           print("n[THOUGHTS GENERATED]n", content material)
    
           return json.hundreds(content material)
    
       besides Exception as e:
    
           print("[Error in LLM Generation]", e)
    
           return []

    This may look just a little lengthy however the important a part of the operate is the immediate half which not solely explains the foundations of the sport to the LLM (find out how to perceive the board, which strikes are legitimate, and so on. ) and likewise the reasoning behind every legitimate transfer. Furthermore, it tells find out how to assign a rating to every potential transfer. These assemble our branches of ideas and at last, our tree ToT. For instance, we’ve a step-by-step information:

    1. Listing {ok} potential cells to click on subsequent.
    
    2. For every, clarify why it may be secure (primarily based on adjoining numbers and identified information).
    
    3. Charge every transfer from 0.0 to 1.0 as a security rating (1 = positively secure).

    These strains information the LLM to suggest a number of strikes and to justify every of those strikes primarily based on the present state; it then has to judge every of those potential strikes by a rating starting from 0 to 1. The agent will use these scores to search out the most suitable choice.

    We now construct an LLM agent utilizing these generated ideas to maneuver a “actual” transfer. Take into account the next code:

    def tot_llm_agent(board, revealed, flagged_mines, known_safe):
    
       ideas = llm_generate_thoughts(board, revealed, flagged_mines, known_safe, ok=5)
    
       if not ideas:
    
           print("[ToT] Falling again to baseline agent as a consequence of no ideas.")
    
           return baseline_agent(board, revealed)
    
       ideas = [t for t in thoughts if 0 <= t["cell"][0] < BOARD_SIZE and 0 <= t["cell"][1] < BOARD_SIZE]
    
       ideas.type(key=lambda x: x["score"], reverse=True)
    
       for t in ideas:
    
           if t["score"] >= 0.9:
    
               transfer = tuple(t["cell"])
    
               print(f"[ToT] Confidently selecting {transfer} with rating {t['score']}")
    
               return transfer
    
       print("[ToT] No high-confidence transfer discovered, utilizing baseline.")
    
       return baseline_agent(board, revealed)

    The agent first calls the LLM to recommend a number of potential subsequent strikes with the boldness rating. If the LLM fails to return any thought, the agent will fall again to a baseline agent outlined earlier and it will probably solely make random strikes. If we’re lucky sufficient to get a number of strikes proposed by the LLM, the agent will don a primary filter to exclude invalid strikes such these which fall out of the board. It can then type the legitimate ideas in accordance with the boldness rating in a descending order and returns one of the best transfer if the rating is larger than 0.9. If not one of the ideas are larger than this threshold, it falls again to the baseline agent.

    Play

    We are going to now attempt to play an ordinary 8×8 Minesweeper board recreation with 10 hidden mines. We performed 10 video games and reached an accuracy of 100%! Please test the notebook for full codes.

    Conclusion

    ToT prompting offers LLMs similar to GPT-4o extra reasoning potential, going past quick and intuitive pondering. We’ve utilized ToT to the Minesweeper recreation and obtained good outcomes. This instance reveals that the ToT can rework LLMs from chat assistants to sophisticated drawback solvers with actual logic and reasoning potential.



    Source link

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Editor Times Featured
    • Website

    Related Posts

    Loop Engineering for RAG Question Parsing: The Small Loop That Runs Before Retrieval

    July 19, 2026

    How to Find the Optimal Coding Agent Interface

    July 9, 2026

    I Completed Five Years in Analytics Consulting: 5 Lessons That Changed How I Work

    June 29, 2026

    GPU-Resident Top-K for Agentic RAG: I Built a CUDA Kernel So My Retrieval Step Would Stop Bouncing Off the GPU

    June 19, 2026

    Can Machine Learning Predict the World Cup?

    June 9, 2026

    Automate Writing Your LLM Prompts

    June 5, 2026

    Comments are closed.

    Editors Picks

    These Were My Favorite Things Samsung Unpacked During Its 2026 Galaxy Event

    July 22, 2026

    AI minister role boosted but tech department axed in Burnham shake-up

    July 21, 2026

    Loop Engineering for RAG Question Parsing: The Small Loop That Runs Before Retrieval

    July 19, 2026

    The risk of weather data sabotage is rising

    July 18, 2026
    Categories
    • Founders
    • Startups
    • Technology
    • Profiles
    • Entrepreneurs
    • Leaders
    • Students
    • VC Funds
    About Us
    About Us

    Welcome to Times Featured, an AI-driven entrepreneurship growth engine that is transforming the future of work, bridging the digital divide and encouraging younger community inclusion in the 4th Industrial Revolution, and nurturing new market leaders.

    Empowering the growth of profiles, leaders, entrepreneurs businesses, and startups on international landscape.

    Asia-Middle East-Europe-North America-Australia-Africa

    Facebook LinkedIn WhatsApp
    Featured Picks

    Logistics giant WiseTech cuts 2000 coding jobs as AI takes over

    February 25, 2026

    All Gemini Users Can Now Use the Viral Nano Banana AI Image Generator. Here’s How

    October 3, 2025

    AI mania pushes Nvidia to record $4 trillion valuation

    July 9, 2025
    Categories
    • Founders
    • Startups
    • Technology
    • Profiles
    • Entrepreneurs
    • Leaders
    • Students
    • VC Funds
    Copyright © 2024 Timesfeatured.com IP Limited. All Rights.
    • Privacy Policy
    • Disclaimer
    • Terms and Conditions
    • About us
    • Contact us

    Type above and press Enter to search. Press Esc to cancel.