Lukas Schäfer - Decision Making in Modern Video Games From Human Play to World Models
This talk presents recent research on decision-making in modern video games conducted at Microsoft Research Cambridge. After motivating why video games are compelling testbeds for studying decision-making, I will present work examining how the choice of visual encoders affects the performance and training efficiency of behaviour cloning (BC) agents. Our experiments show that carefully selected pre-trained visual encoders can significantly reduce computational cost and boost performance. Given the substantial data requirements of BC for complex tasks, we then investigate predictive inverse dynamics models (PIDM)—which condition a policy on predicted future states—as an alternative to BC. These models have demonstrated improved performance over BC but remain poorly understood. I will present theoretical insights that show that PIDM's performance gains can be explained with a bias-variance tradeoff: conditioning the policy on future context can reduce uncertainty about action predictions but also introduce bias whenever future predictions are inaccurate. We further show that these insights translate into significant sample efficiency gains in 2D navigation tasks and complex 3D environments in modern video games. Finally, we move from decision-making models that model the future to world and human action models (WHAM), which combine an environment model (world model) with an imitation-learning policy representing human gameplay. Inspired by the recipe behind LLMs, we demonstrate the promise of scale for such models and explore how they can support workflows for video game creatives.
Lukas Schäfer is a postdoctoral researcher at Microsoft Research in Cambridge, UK, where he is part of Katja Hofmann's team working on machine learning for video games. His work focuses on developing autonomous agents that enable novel experiences and tools in video games, with an emphasis on imitation learning approaches. Lukas holds a PhD and MSc in Informatics from the University of Edin
Watch on YouTube ↗
(saves to browser)
Sign in to unlock AI tutor explanation · ⚡30
Playlist
Uploads from Cohere · Cohere · 0 of 60
← Previous
Next →
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
48
49
50
51
52
53
54
55
56
57
58
59
60
Andreas Madsen on Independent Research and Interpretability
Cohere
Plex: Towards Reliability using Pretrained Large Model Extensions
Cohere
Independent Research Panel Discussion
Cohere
The Future of ML Ops: Open Challenges and Opportunities
Cohere
C4AI Special - Grad School Applications
Cohere
Cohere For AI Fireside Chat: Samy Bengio
Cohere
Cohere For AI - Scholars Program Information Session
Cohere
Modular and Composable Transfer Learning with Jonas Pfeiffer
Cohere
Jay Alammar Presents Large Language Models for Real World Applications
Cohere
Catherine Olsson - Mechanistic Interpretability: Getting Started
Cohere
How To Prompt Engineer a Tech Interview App | TOHacks 2022 Winners
Cohere
C4AI Sparks: Samy Bengio
Cohere
BERTopic for Topic Modeling - Maarten Grootendorst - Talking Language AI Ep#1
Cohere
Exploring News Headlines With Text Clustering | Jay Alammar
Cohere
Scale TransformX | Fireside Chat: Aidan Gomez and Alexandr Wang
Cohere
Making Large Language Models Accessible | Scale AI Fireside chat with Bill MacCartney
Cohere
Intro to KeyBERT - BERTopic for Topic Modeling
Cohere
Intro to PolyFuzz - BERTopic for Topic Modeling
Cohere
API Design Philosophy - BERTopic for Topic Modeling
Cohere
Code demo of BERTopic - BERTopic for Topic Modeling
Cohere
Short texts vs long texts in BERTopic- BERTopic for Topic Modeling
Cohere
How People can help BERTopic - BERTopic for Topic Modeling
Cohere
Cohere For AI: Training Sensorimotor Agency in Cellular Automata with Bert Chan
Cohere
Cohere API Community Demos | October 2022
Cohere
Perfect Prompt Demo By Arjun Patel
Cohere
Project Idea Generator Demo By Tobechukwu Okamkpa
Cohere
SuperTransformer Demo By Amir Nagri and Team Megatron
Cohere
Cohere For AI Fireside Chat: Pablo Samuel Castro
Cohere
How Startups Can Use NLP to Build a Competitive Moat
Cohere
Build Chatbots Faster with Large Language Models
Cohere
Tools to Improve Training Data - Vincent Warmerdam - Talking Language AI Ep#2
Cohere
Utku Evci - Sparsity and Beyond Static Network Architectures
Cohere
Adding human intelligence to ML models with human-learn #shorts #machinelearning #nlp
Cohere
Iterating on your data with doubtlab - Tools to Improve Training Data
Cohere
Adding Human Intelligence to ML models with Human learn - Tools to Improve Training Data
Cohere
Scikt Learn embeddings helpers with Embetter - Tools to Improve Training Data
Cohere
Building Cohere API Demo App With Streamlit | Adrien Morisot
Cohere
Rosanne Liu - career creation for non-standard candidates
Cohere
Giving computers many human languages with Cohere's multilingual embeddings
Cohere
Learning by Distilling Context with Charlie Snell
Cohere
Sentence Transformers and Embedding Evaluation - Nils Reimers - Talking Language AI Ep#3
Cohere
Reflecting on for.ai...
Cohere
Create a Custom Language Model with Surge AI and Cohere
Cohere
Cohere API Community Demos | November 2022
Cohere
Cohere API Community Demos | December 2022
Cohere
Cohere For AI Presents: Colin Raffel
Cohere
Lucas Beyer - FlexiViT: One Model for All Patch Sizes
Cohere
What is Neural Search? Nils Reimers - Sentence Transformers and Embedding Evaluation
Cohere
Evaluating Information Retrieval with BEIR
Cohere
Evaluating Embeddings with MTEB Massive text embeddings benchmark - Nils Reimers
Cohere
High quality text classification with few training examples with SetFit
Cohere
Multilingual and cross lingual embeddings - Nils Reimers
Cohere
Developing open-source software: lessons, benefits, and challenges - Nils Reimers
Cohere
Ask Me Anything with Ed Grefenstette, Head of Machine Learning at Cohere
Cohere
HyperWrite Powers Its Generative AI Service with Cohere
Cohere
EMNLP 2022 Conference Special Edition - Talking Language AI #4
Cohere
Cohere API Community Demos | January 2023
Cohere
C4AI Sparks: Rosanne Liu on Career Creation for Non-Standard Candidates
Cohere
Michael Tschannen - Image-and-Language Understanding from Pixels Only
Cohere
How to Add AI to your App
Cohere
More on: Agent Foundations
View skill →Related AI Lessons
⚡
⚡
⚡
⚡
The Future of Human Creativity in the Age of AI
Medium · AI
The Boring “Multi-Agent” Loop That Quietly Earns $2,000/Month (With Zero Maintenance)
Medium · Machine Learning
Building OMEGA — A Cinematic Multi-Agent IPL Strategy Engine Powered by Google Gemini
Dev.to · Omkar Rane
🏏 Captain Cool — Orchestrating a Google Gemini Multi-Agent Debate Loop for Live IPL Strategy
Dev.to · siddhi bhosale
🎓
Tutor Explanation
DeepCamp AI