ia generativa
as raízes em 2014 e o caminho de 2017 ao agora: do transformer aos agentes, na ordem em que aconteceu.
- Efficient Estimation of Word Representations in Vector Space
- Neural Machine Translation by Jointly Learning to Align and Translate
- Sequence to Sequence Learning with Neural Networks
- Neural Machine Translation of Rare Words with Subword Units
- Attention Is All You Need
- Deep Reinforcement Learning from Human Preferences
- Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer
- BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
- Improving Language Understanding by Generative Pre-Training
- Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
- Language Models are Unsupervised Multitask Learners
- An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
- Denoising Diffusion Probabilistic Models
- Language Models are Few-Shot Learners
- Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks
- Scaling Laws for Neural Language Models
- Evaluating Large Language Models Trained on Code
- Finetuned Language Models Are Zero-Shot Learners
- High-Resolution Image Synthesis with Latent Diffusion Models
- Learning Transferable Visual Models From Natural Language Supervision
- LoRA: Low-Rank Adaptation of Large Language Models
- RoFormer: Enhanced Transformer with Rotary Position Embedding
- Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity
- Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
- Constitutional AI: Harmlessness from AI Feedback
- Emergent Abilities of Large Language Models
- Fast Inference from Transformers via Speculative Decoding
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness
- PaLM: Scaling Language Modeling with Pathways
- ReAct: Synergizing Reasoning and Acting in Language Models
- Self-Consistency Improves Chain of Thought Reasoning in Language Models
- Self-Instruct: Aligning Language Models with Self-Generated Instructions
- Training Compute-Optimal Large Language Models
- Training language models to follow instructions with human feedback
- Are Emergent Abilities of Large Language Models a Mirage?
- Direct Preference Optimization: Your Language Model is Secretly a Reward Model
- Efficient Memory Management for Large Language Model Serving with PagedAttention
- Efficient Streaming Language Models with Attention Sinks
- GPT-4 Technical Report
- GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints
- Llama 2: Open Foundation and Fine-Tuned Chat Models
- LLaMA: Open and Efficient Foundation Language Models
- Lost in the Middle: How Language Models Use Long Contexts
- Mamba: Linear-Time Sequence Modeling with Selective State Spaces
- QLoRA: Efficient Finetuning of Quantized LLMs
- Reflexion: Language Agents with Verbal Reinforcement Learning
- SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
- Textbooks Are All You Need
- Toolformer: Language Models Can Teach Themselves to Use Tools
- Tree of Thoughts: Deliberate Problem Solving with Large Language Models
- Alignment Faking in Large Language Models
- Automated Unit Test Improvement using Large Language Models at Meta
- DeepSeek-V3 Technical Report
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
- Mixtral of Experts
- Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
- Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet
- Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training
- The Llama 3 Herd of Models
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- On the Biology of a Large Language Model