Skip to yearly menu bar Skip to main content


Poster Wed, Dec 9, 2026 • 10:00 AM – 1:00 PM AEDT Hall 1-4

Self-Supervised On-Policy Reinforcement Learning via Contrastive Proximal Policy Optimisation

Asim Osman ⋅ Sasha Abramowitz ⋅ Mark Bergh ⋅ Ulrich Armel Mbou Sob ⋅ Ruan John de Kock ⋅ Omayma Mahjoub ⋅ Oussama Hidaoui ⋅ Noah De Nicola ⋅ Arnol M Fokam ⋅ Felix Chalumeau ⋅ Daniel Rajaonarivonivelomanantsoa ⋅ Siddarth Singh ⋅ Refiloe Shabe ⋅ Juan Formanek ⋅ Simon Du Toit ⋅ Arnu Pretorius

Abstract

Chat is not available.