MIT Researchers Enhanced Artificial Intelligence (AI) 64x Better at Planning, Achieving 94% Accuracy
Can a 8B-parameter language mannequin produce provably legitimate multi-step plans as an alternative of believable guesses? MIT CSAIL researchers introduce PDDL-INSTRUCT, an instruction-tuning framework that {couples} logical chain-of-thought with exterior plan validation (VAL) to carry symbolic planning efficiency of LLMs. On PlanBench, a tuned Llama-3-8B reaches 94% legitimate plans on Blocksworld, with massive jumps on…
