Pau Labarta Bajo's Newsletter

Pau Labarta Bajo's Newsletter

Fine-tuning a Small Language Model for browser control with GRPO and OpenEnv

Pau Labarta Bajo's avatar
Pau Labarta Bajo
Dec 27, 2025
∙ Paid

Today I want to share with you the write-up of a live 60-minute session I hosted on the Liquid AI Discord Community.

The topic? How to teach Language Models to navigate websites and complete tasks using Reinforcement Learning.

We’re talking about building browser agents that can click buttons, fill forms, and even book flights, all by learning from trial …

This post is for subscribers in the Founding Member plan

Already in the Founding Member plan? Sign in
© 2026 Pau Labarta Bajo · Privacy ∙ Terms ∙ Collection notice
Start your SubstackGet the app
Substack is the home for great culture