Statistical inference of additive monotone models and instance-optimality of multi-armed bandits and reinforcement learning | Synapse