Sources
fly51fly[LG] LLM-AutoDiff: Auto-Differentiate Any LLM Workflow L Yin, Z W (Atlas) [SylphAI & University of Texas at Austin] (2025) https://t.co/8ywZtJHhX8 https://t.co/jNVAX4zsFr
fly51fly[LG] Towards General-Purpose Model-Free Reinforcement Learning S Fujimoto, P D'Oro, A Zhang, Y Tian... [Meta] (2025) https://t.co/cg4gVMObol https://t.co/6zf5Q0MEBb
fly51fly[LG] Improving Your Model Ranking on Chatbot Arena by Vote Rigging R Min, T Pang, C Du, Q Liu... [Sea AI Lab] (2025) https://t.co/ICsX9Sxctm https://t.co/YwLY5zcDAi
Additional media

























