AI News selected for Professionals and Decision Makers
Primary Research Stream

Large Behavior Model: A Promptable Digital Twin of the Retail Customer

06:00 · July 9, 2026 · arXiv cs.AI RSS

Large Behavior Model: A Promptable Digital Twin of the Retail Customer

Customer behavior modeling underpins recommendation, marketing, and decision support, yet existing approaches either optimize predictive accuracy without explaining decisions or simulate users without grounding them in real behavioral data. We present the Large Behavioral Model (LBM) that learns customer decision making directly from large-scale retail transactions through a unified Person-Environment formulation. Customer state is represented by a behavioral profile derived from historical purchases, while product context is incorporated through retrieval-augmented generation. The model is trained using continued pre-training on verbalized behavioral data, supervised fine-tuning for decision generation, and reinforcement learning with verifiable rewards for evidence-based calibration. We evaluate the proposed framework on purchase prediction, hard-negative discrimination, basket completion, promotion response, and cross-domain voucher redemption. The model consistently outperforms frontier general-purpose language models on in-domain retail tasks while demonstrating strong zero-shot and fine-tuned transfer across retailers and decision domains. Ablation studies show that continued pre-training is the primary driver of behavioral generalization, retrieval is most effective when applied during both training and inference, and reinforcement learning improves reliance on explicit behavioral evidence over generic language-model priors. These results demonstrate that behavioral knowledge encoded in transaction histories can be effectively learned by language models, providing a scalable foundation for customer digital twins and behavior simulation.

Summary

The Large Behavioral Model (LBM) addresses shortcomings in existing customer modeling techniques, which typically prioritize predictive accuracy at the expense of interpretability or generate synthetic user simulations detached from actual transaction histories. The approach frames customer decision making through a unified Person-Environment formulation in which a behavioral profile, derived directly from historical purchases, represents the customer state while product context is supplied via retrieval-augmented generation.

Training proceeds in three stages. Continued pre-training on verbalized behavioral sequences allows the model to internalize patterns latent in large-scale retail data. Supervised fine-tuning then aligns outputs with decision-generation objectives, after which reinforcement learning with verifiable rewards calibrates the model to favor explicit evidence drawn from the retrieved context rather than generic language-model priors.

Evaluations span purchase prediction, hard-negative discrimination, basket completion, promotion response, and cross-domain voucher redemption. Across these tasks the LBM outperforms frontier general-purpose language models on in-domain retail benchmarks and exhibits robust zero-shot and fine-tuned transfer to new retailers and decision domains. Ablation results indicate that continued pre-training contributes most to behavioral generalization, retrieval yields the largest gains when applied at both training and inference time, and reinforcement learning strengthens reliance on observable transaction evidence.

Why it matters

This research is highly relevant for Dutch AI researchers and practitioners in the robust local retail and e-commerce sectors (e.g., Bol.com, Ahold Delhaize). The methodology offers an actionable, transparent approach to customer modeling that aligns with the EU's demand for explainable and evidence-based AI systems.

More in this beat
digital-twinlarge-behavioral-modellarge-language-modelsrecommender-systemsreinforcement-learningretrieval-augmented-generationsupervised-fine-tuning
Beyond Trajectory Imitation: Strategy-Guided Policy Optimization for LLM Reasoning

06:00 · June 24, 2026

Beyond Trajectory Imitation: Strategy-Guided Policy Optimization for LLM Reasoning

This research provides advanced methodologies for LLM distillation, which is crucial for Dutch AI researchers aiming to develop efficient, high-performing local models. The shift from memorization to strategy acquisition aligns with the Netherlands' focus on robust, generalizable, and sustainable AI systems.

Relevance 85 · Audience 95

TAPR: Enhancing LLM Performance with a Task-Aware Prompt Rewriter

06:00 · August 3, 2026

TAPR: Enhancing LLM Performance with a Task-Aware Prompt Rewriter

Directly applicable by Dutch AI teams via public code; strong technical depth and novelty in prompt optimization using GRPO and LLM judges; Dutch institutional ties (UvA) and relevance to EU LLM deployment and ethical AI practices.

Relevance 82 · Audience 88

Aligning Clinical Needs and AI Capabilities: A Survey on LLMs for Medical Reasoning

06:00 · July 11, 2026

Aligning Clinical Needs and AI Capabilities: A Survey on LLMs for Medical Reasoning

This survey provides a rigorous, structured framework for evaluating medical LLMs, which is highly valuable for Dutch AI researchers and healthcare institutions developing transparent and safe clinical AI. Its focus on mitigating hallucinations and ensuring reliable reasoning aligns well with the EU AI Act and the Netherlands' emphasis on ethical AI deployment.

Relevance 85 · Audience 95

Agentic AI and Retrieval-Augmented Models in Straight-Through Underwriting

06:00 · July 11, 2026

Agentic AI and Retrieval-Augmented Models in Straight-Through Underwriting

The article is highly relevant for Dutch AI researchers and InsurTech practitioners as it provides a concrete, reproducible framework for deploying multi-agent LLM systems in highly regulated domains. Its strong emphasis on auditability, transparency, and human-in-the-loop governance aligns perfectly with the EU AI Act and the Netherlands' strategic focus on ethical AI.

Relevance 85 · Audience 95

Human-Centric Reflective Architecture for Human-AI Collaborative Decision-Making

06:00 · July 7, 2026

Human-Centric Reflective Architecture for Human-AI Collaborative Decision-Making

This research is highly relevant to the Dutch AI market's focus on ethical, transparent, and human-centric AI. The proposed HCRA framework provides advanced methodologies for researchers to build AI systems that align with human preferences, directly supporting EU AI Act compliance regarding human oversight.

Relevance 85 · Audience 95

Reinforcement Learning for Evidence-Seeking Diagnostic Reasoning with Large Language Models

06:00 · July 7, 2026

Reinforcement Learning for Evidence-Seeking Diagnostic Reasoning with Large Language Models

This research is highly relevant for Dutch AI researchers and health-tech enterprises developing autonomous clinical assistants. The use of RLVR and RAGES provides a novel, actionable methodology for creating more accurate, iterative, and verifiable medical AI systems, aligning with the EU's focus on robust healthcare AI.

Relevance 85 · Audience 95

Oyster-II: Reinforcement Learning for Constructive Safety Alignment in Large Language Models

06:00 · July 7, 2026

Oyster-II: Reinforcement Learning for Constructive Safety Alignment in Large Language Models

This research is highly relevant to the Dutch AI market due to the Netherlands' and EU's strong regulatory focus on ethical, safe, and transparent AI. Oyster-II provides advanced researchers with actionable RL methodologies to align LLMs safely without compromising their utility, directly supporting compliant AI development.

Relevance 85 · Audience 95