RLCD: Reinforcement Learning for Calibrated Decisions
How RLCD uses proper scoring rules and calibration penalties to stop LLMs from bluffing.
How RLCD uses proper scoring rules and calibration penalties to stop LLMs from bluffing.
Image Classification Model and Web App using Flask.
A Breakdown of the Transformer Model.
Semantic Segmentation of Satellite Images.