Live
🔥 google-research/timesfmGitHub Trending🔥 aliasrobotics/caiGitHub Trending🔥 ComposioHQ/awesome-claude-skillsGitHub Trending🔥 SkyworkAI/Matrix-GameGitHub Trending🔥 sponsors/vas3kGitHub Trending🔥 sponsors/khoj-aiGitHub Trending🔥 PaddlePaddle/PaddleOCRGitHub TrendingTest: 15% of Americans say they would work for AI bossTechCrunch AIAutoMS: Multi-Agent Evolutionary Search for Cross-Physics Inverse Microstructure DesignarXivMultiverse: Language-Conditioned Multi-Game Level Blending via Shared RepresentationarXivMediHive: A Decentralized Agent Collective for Medical ReasoningarXivBitboard version of Tetris AIarXivThe Price of Meaning: Why Every Semantic Memory System ForgetsarXivWhen Verification Hurts: Asymmetric Effects of Multi-Agent Feedback in Logic Proof TutoringarXivQuantification of Credal Uncertainty: A Distance-Based ApproacharXiv🔥 google-research/timesfmGitHub Trending🔥 aliasrobotics/caiGitHub Trending🔥 ComposioHQ/awesome-claude-skillsGitHub Trending🔥 SkyworkAI/Matrix-GameGitHub Trending🔥 sponsors/vas3kGitHub Trending🔥 sponsors/khoj-aiGitHub Trending🔥 PaddlePaddle/PaddleOCRGitHub TrendingTest: 15% of Americans say they would work for AI bossTechCrunch AIAutoMS: Multi-Agent Evolutionary Search for Cross-Physics Inverse Microstructure DesignarXivMultiverse: Language-Conditioned Multi-Game Level Blending via Shared RepresentationarXivMediHive: A Decentralized Agent Collective for Medical ReasoningarXivBitboard version of Tetris AIarXivThe Price of Meaning: Why Every Semantic Memory System ForgetsarXivWhen Verification Hurts: Asymmetric Effects of Multi-Agent Feedback in Logic Proof TutoringarXivQuantification of Credal Uncertainty: A Distance-Based ApproacharXiv

Bitboard version of Tetris AI

arXivby [Submitted on 24 Mar 2026]March 31, 20262 min read2 views
Source Quiz

arXiv:2603.26765v1 Announce Type: new Abstract: The efficiency of game engines and policy optimization algorithms is crucial for training reinforcement learning (RL) agents in complex sequential decision-making tasks, such as Tetris. Existing Tetris implementations suffer from low simulation speeds, suboptimal state evaluation, and inefficient training paradigms, limiting their utility for large-scale RL research. To address these limitations, this paper proposes a high-performance Tetris AI framework based on bitboard optimization and improved RL algorithms. First, we redesign the Tetris game — Xingguo Chen, Pingshou Xiong, Zhenyu Luo, Mengfei Hu, Xinwen Li, Yongzhou L\"u, Guang Yang, Chao Li, Shangdong Yang

View PDF HTML (experimental)

Abstract:The efficiency of game engines and policy optimization algorithms is crucial for training reinforcement learning (RL) agents in complex sequential decision-making tasks, such as Tetris. Existing Tetris implementations suffer from low simulation speeds, suboptimal state evaluation, and inefficient training paradigms, limiting their utility for large-scale RL research. To address these limitations, this paper proposes a high-performance Tetris AI framework based on bitboard optimization and improved RL algorithms. First, we redesign the Tetris game board and tetrominoes using bitboard representations, leveraging bitwise operations to accelerate core processes (e.g., collision detection, line clearing, and Dellacherie-Thiery Features extraction) and achieve a 53-fold speedup compared to OpenAI Gym-Tetris. Second, we introduce an afterstate-evaluating actor network that simplifies state value estimation by leveraging Tetris afterstate property, outperforming traditional action-value networks with fewer parameters. Third, we propose a buffer-optimized Proximal Policy Optimization (PPO) algorithm that balances sampling and update efficiency, achieving an average score of 3,829 on 10x10 grids within 3 minutes. Additionally, we develop a Python-Java interface compliant with the OpenAI Gym standard, enabling seamless integration with modern RL frameworks. Experimental results demonstrate that our framework enhances Tetris's utility as an RL benchmark by bridging low-level bitboard optimizations with high-level AI strategies, providing a sample-efficient and computationally lightweight solution for scalable sequential decision-making research.

Subjects:

Artificial Intelligence (cs.AI)

Cite as: arXiv:2603.26765 [cs.AI]

(or arXiv:2603.26765v1 [cs.AI] for this version)

https://doi.org/10.48550/arXiv.2603.26765

arXiv-issued DOI via DataCite (pending registration)

Submission history

From: Xingguo Chen [view email] [v1] Tue, 24 Mar 2026 02:35:09 UTC (1,254 KB)

Original source

arXiv

Was this article helpful?

Sign in to highlight and annotate this article

AI
Ask AI about this article
Powered by AI News Hub · full article context loaded
Ready

Conversation starters

Ask anything about this article…

Daily AI Digest

Get the top 5 AI stories delivered to your inbox every morning.

Knowledge Map

Knowledge Map
TopicsEntitiesSource
Bitboard ve…researchpaperarxivaiartificial-…arXiv

Connected Articles — Knowledge Graph

This article is connected to other articles through shared AI topics and tags.

Knowledge Graph100 articles · 336 connections
Scroll to zoom · drag to pan · click to open

Discussion

Sign in to join the discussion

No comments yet — be the first to share your thoughts!

More in Research Papers