Imagine a nightclub bouncer who doesn't care what music is playing inside — their only job is to stop anyone from going through the wrong door. They don't change the DJ's playlist, they don't move the furniture, they just redirect people at the threshold. That's FAITH: a learned safety filter that sits between a task-trained RL policy and the actuators, rewriting only the dangerous actions and leaving everything else untouched. The DJ (task policy) never has to think about safety; the bouncer (filter) never has to think about music. The core claim is architectural, not algorithmic: separating safety from task optimization at the action level, using a learned (not analytic) safety value function, eliminates the gradient tug-of-war that plagues constrained RL. Standard safe RL folds a safety penalty into the reward, creating competing gradients — the policy simultaneously tries to walk forward and avoid obstacles, and the optimizer compromises on both. FAITH's task policy sees only task reward; the filter intercepts unsafe actions post-hoc. This is not a new idea in principle — Hamilton-Jacobi safety filters do this with analytic dynamics — but FAITH does it model-free, using a feedforward network trained on the learned safety Q-function, making it applicable where you don't have a dynamics model. The technical novelty is the feasibility-aware fallback. Classical minimal-intervention filters assume there is always a safe action — they project onto the safe set. In high-dimensional systems, the safe set can be empty from certain states (the humanoid is already falling). FAITH handles this gracefully: when no action satisfies the learned safety condition, the filter outputs the action with the minimum predicted peak harm. This is a genuine design contribution, because hard projections are undefined in infeasible states, and most prior work simply crashes or reverts to a default. Results span three environments of increasing difficulty. On a double integrator and Safety Gym's PointGoal, FAITH achieves the highest task return among methods with zero feasible-start safety violations and matches the lowest harm from infeasible starts. The headline number is the 29-DoF Unitree H1 humanoid in MuJoCo: 99.95% safety rate on Walking-Avoid with 97% of unconstrained return preserved. In Push-Avoid, the humanoid learns to deliberately fall away from the protected region — sacrificing balance to guarantee safety — which is a qualitatively interesting emergent behavior the authors highlight. The ladder comparison is honest but limited. FAITH is benchmarked against PPO-Lagrangian, SAC-Lagrangian, SQRL, and Recovery RL on Safety Gym, plus unconstrained PPO and PPO-Lagrangian on the humanoid tasks. These are the standard safe RL baselines, but the field has been moving: recent model-based safety filter work (e.g., MBSF, SafeDreamer) and diffusion-based planners are not compared. The humanoid demonstration is impressive but the comparison set is thin — only unconstrained PPO and PPO-Lagrangian, not the full battery of safe RL methods adapted to high-DoF locomotion. Integrity is mixed. The simulation results are well-ablated and the real-robot demonstration on a Unitree G1 adds credibility, but there's no independent replication, no pre-registration, and the benchmark environments (Safety Gym, MuJoCo humanoid) were likely chosen because they favor the method's strengths. The real-world transfer is demonstrated but not quantified with the same metrics — it's a qualitative 'it works' demonstration rather than a systematic sim-to-real evaluation. The obvious next experiment is scaling to contact-rich manipulation and outdoor locomotion with unmodeled disturbances. The safety Q-function is trained in simulation — how it degrades under distribution shift is the open question. The authors likely didn't run this because the sim-to-real gap for safety guarantees requires either robust training or formal verification, neither of which is cheap. The feasibility-aware fallback is the most transferable idea here: expect to see it adopted independently of the rest of the framework.