![I Trained an AI to Beat Final Fight… Here’s What Happened [p]](https://external-preview.redd.it/dWZSKa_lMUvycB0q8xwsIkTgDpHLe-W2-Q_S7RwWucQ.jpeg?width=320&crop=smart&auto=webp&s=6a0cfa02507091d2949c8f6b2fcd59a254a23929)
Machine Learning
I Trained an AI to Beat Final Fight… Here’s What Happened [p]
In this post, I delve into my experience training an AI agent using Behavior Cloning on the classic arcade game Final Fight. By relying solely on demonstrations, I evaluated the agent's performance in the first stage, navigating challenges like action space remapping and trajectory alignment issues. While the agent shows promise, consistency and survival remain hurdles. I’m eager for community insights on enhancing BC performance, transitioning to PPO, and addressing partial observability.
![torch-nvenc-compress: GPU NVENC silicon as a PCIe bandwidth multiplier — PCA + pure-ctypes Video Codec SDK wrapper. Parallel-path overlap measured at 67% of theoretical max on a real GEMM + encode workload. [P]](https://external-preview.redd.it/vqLrMLU0urgSqpiud1c7Ilq7WSsJhRPX63HDDrDRN6M.png?width=640&crop=smart&auto=webp&s=0d43a15121928a0c4b5e3a9730e67ff06df77324)












