
First-Person Fairness in Chatbots
⚖️ First-Person Fairness in ChatbotsThis paper from OpenAI examines potential bias in chatbot systems like ChatGPT, specifically focusing on how a user's name, which can be associated with demographic...
18 Loka 20249min

Thinking LLMs
🤔 Thinking LLMs: General Instruction Following with Thought GenerationThis research paper explores the concept of "Thinking LLMs," or large language models that can generate internal thoughts before ...
18 Loka 202419min

Addition is All You Need
🔋 Addition is All You Need for Energy-efficient Language ModelsThis research paper introduces a novel algorithm called Linear-Complexity Multiplication (L-Mul) that aims to make language models more ...
18 Loka 20249min

MLE-bench
🤖 MLE-bench: Evaluating Machine Learning Agents on Machine Learning EngineeringThe paper introduces MLE-bench, a benchmark designed to evaluate AI agents' ability to perform machine learning engineer...
18 Loka 202412min

Long-Context LLMs Meet RAG
📈 Long-Context LLMs Meet RAG: Overcoming Challenges for Long Inputs in RAGThis paper explores the challenges and opportunities of using long-context language models (LLMs) in retrieval-augmented gene...
18 Loka 202415min

GSM-Symbolic
📊 GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language ModelsThe paper investigates the mathematical reasoning abilities of large language models (LLMs). The author...
18 Loka 20246min

Anti-Social LLM
😶 Anti-Social Behavior and Persuasion Ability of LLMsThis study explores the behavior of Large Language Models (LLMs) in a simulated prison environment, inspired by the Stanford Prison Experiment. It...
18 Loka 20248min

Differential Transformer
🎧 Differential TransformerThe paper introduces the Differential Transformer, a new architecture for large language models (LLMs) that aims to improve their ability to focus on relevant information wi...
18 Loka 20249min



















