TY - RPRT TI - Per-Domain Generalizing Policies: On Learning Efficient and Robust Q-Value Functions (Extended Version with Technical Appendix) AU - Nicola J. Müller AU - Moritz Oster AU - Isabel Valera AU - Jörg Hoffmann AU - Timo P. Gros PY - 2026 UR - https://arxiv.org/abs/2603.17544 ID - 2603.17544 ER -