








Lightweight Guide to understanding GRPO and RL principles
A beginner-friendly guide to Group Relative Policy Optimization (GRPO) training workflow without assuming prior RL knowledge.

Having no Strategy is a Strategy in Ottawa
What I said at my recent keynote for the Ottawa Tech Investment Summit — and what the government said three days later.

Playing to Win: How Strategy Really Works
How Strategy Really Works. This approach grew out of the strategy practice at Monitor Company and subsequently became the standard process at P&G.

A Strategy and Wardley Mapping Primer
If you're doing strategic work, think in terms of evolutionary flow — and learn to draw fast-and-dirty Wardley maps.

Strategy in four worlds
Different environments require different survival strategies

The long game: China's grand strategy to displace American order | Brookings
What are China’s ambitions, and does it have a grand strategy to achieve them? If it does, what is that strategy, what shapes it, and what should the United States do about it? Rush Doshi attempts to provide an answer in his new book, "The Long Game."

The French nuclear deterrent in a changing strategic environment | Note de la FRS | Foundation for Strategic Research | FRS
The Foundation for Strategic Research (FRS) is France's main centre of expertise on international security and defence issues.

The Strategic Foresight Book — IFRC Solferino Academy
Ready to shape the future? Download the Foresight Book This book will help you engage with uncertainty, navigate complex challenges, and build more resilient organisations and communities.

Strategic interdependence is rewiring the global economy
It is no longer sufficient for the US-China trade relationship to be driven solely by cost and efficiency

#Exploration: A Study of Count-Based Exploration for Deep...
Count-based exploration algorithms are known to perform near-optimally when used in conjunction with tabular reinforcement learning (RL) methods for solving small discrete Markov decision...

Memory Efficient RL | Unsloth Documentation
We're excited to introduce more efficient reinforcement learning (RL) in Unsloth with multiple algorithmic advancements:

Is the map of global bat distribution in this paper (Nov 2025, ‘Sustainability’ journal) AI nonsense? Comments from experts on LinkedIn are suggesting so. mdpi.com/2071-1050/17/22/10339 LinkedIn: linkedin.com/posts/alice-hughes-509b45148_…