论文库
论文与网络文章:研究时的一手参考。读完我们的拆解,不必回去看原文。
收 139 篇:论文 139 篇、网络文章 0 篇; 其中 0 篇已经拆完。
判据和书架一样:你读完我们的拆解,不必回去看原文。
这些论文是怎么进来的
第一批 139 篇,来自书架自己的引用 —— 85 本书的拆解里引到它们、却一篇都没读过原文。先补这个缺口。
怎么引用它
paper=<id>@arXiv:<编号><版本> §<节> 论文,版本号必须钉死
article=<id>@<抓取日期> §<小节> 文章,抓取日期必须钉死
别的书架引这里,用跨库锚:shelf=ai-paper-reference/<id>#index.md。
pretraining —— 预训练、规模律、数据配方(46 篇)
- Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks — 书架上 9 本书引过它
- LoRA: Low-Rank Adaptation of Large Language Models — 书架上 7 本书引过它
- Training Compute-Optimal Large Language Models — 书架上 7 本书引过它
- Language Models are Few-Shot Learners — 书架上 6 本书引过它
- Large Language Models are Zero-Shot Reasoners — 书架上 4 本书引过它
- QLoRA: Efficient Finetuning of Quantized LLMs — 书架上 4 本书引过它
- An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale — 书架上 3 本书引过它
- From Local to Global: A Graph RAG Approach to Query-Focused Summarization — 书架上 3 本书引过它
- Self-Consistency Improves Chain of Thought Reasoning in Language Models — 书架上 3 本书引 过它
- BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding — 书架上 2 本书引过它
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models — 书架上 2 本书引过它
- Dense X Retrieval: What Retrieval Granularity Should We Use? — 书架上 2 本书引过它
- Learning Transferable Visual Models From Natural Language Supervision — 书架上 2 本书引过它
- Parameter-Efficient Transfer Learning for NLP — 书架上 2 本书引过它
- Scaling Laws for Neural Language Models — 书架上 2 本书引过它
- ALBERT: A Lite BERT for Self-supervised Learning of Language Representations — 书架上 1 本书引过它
- BERTopic: Neural topic modeling with a class-based TF-IDF procedure — 书架上 1 本书引过它
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models — 书架上 1 本书引过它
- Conditional Prompt Learning for Vision-Language Models — 书架上 1 本书引过它
- Constitutional AI: Harmlessness from AI Feedback — 书架上 1 本书引过它
- DeepSeek-V3 Technical Report — 书架上 1 本书引过它
- Docling Technical Report — 书架上 1 本书引过它
- Efficient Few-Shot Learning Without Prompts — 书架上 1 本书引过它
- Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer — 书架上 1 本书引过它
- FineScope : SAE-guided Data Selection Enables Domain Specific LLM Pruning and Finetuning — 书架上 1 本书引过它
- Generated Knowledge Prompting for Commonsense Reasoning — 书架上 1 本书引过它
- GLM: General Language Model Pretraining with Autoregressive Blank Infilling — 书架上 1 本书引过它
- Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation — 书架上 1 本书引过它
- GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints — 书架上 1 本书引过它
- HellaSwag: Can a Machine Really Finish Your Sentence? — 书架上 1 本书引过它
- Intrinsic Dimensionality Explains the Effectiveness of Language Model Fine-Tuning — 书架上 1 本书引过它
- Learning to Prompt for Vision-Language Models — 书架上 1 本书引过它
- Llama 2: Open Foundation and Fine-Tuned Chat Models — 书架上 1 本书引过它
- Measuring Mathematical Problem Solving With the MATH Dataset — 书架上 1 本书引过它
- Neural Machine Translation of Rare Words with Subword Units — 书架上 1 本书引过它
- Prefix-Tuning: Optimizing Continuous Prompts for Generation — 书架上 1 本书引过它
- Qwen2.5 Technical Report — 书架上 1 本书引过它
- Qwen3 Technical Report — 书架上 1 本书引过它
- RLPrompt: Optimizing Discrete Text Prompts with Reinforcement Learning — 书架上 1 本书引过它
- RoBERTa: A Robustly Optimized BERT Pretraining Approach — 书架上 1 本书引过它
- Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks — 书架上 1 本书引过它
- SentencePiece: A simple and language independent subword tokenizer and detokenizer for Neural Text Processing — 书架上 1 本书引过它
- Show Your Work: Scratchpads for Intermediate Computation with Language Models — 书架上 1 本书引过它
- Subword Regularization: Improving Neural Network Translation Models with Multiple Subword Candidates — 书架上 1 本书引过它
- Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge — 书架上 1 本书引过它
- Understanding R1-Zero-Like Training: A Critical Perspective — 书架上 1 本书引过它
prompting —— 提示工程、思维链、自洽、上下文学习(30 篇)
- Chain-of-Thought Prompting Elicits Reasoning in Large Language Models — 书架上 6 本书引过它
- Language Models are Few-Shot Learners — 书架上 6 本书引过它
- ReAct: Synergizing Reasoning and Acting in Language Models — 书架上 6 本书引过它
- Training language models to follow instructions with human feedback — 书架上 6 本书引过它
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning — 书架上 5 本书引过它
- Large Language Models are Zero-Shot Reasoners — 书架上 4 本书引过它
- Self-Consistency Improves Chain of Thought Reasoning in Language Models — 书架上 3 本书引过它
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models — 书架上 2 本书引过它
- Take a Step Back: Evoking Reasoning via Abstraction in Large Language Models — 书架上 2 本书引过它
- Tree of Thoughts: Deliberate Problem Solving with Large Language Models — 书架上 2 本书引过它
- Automatic Chain of Thought Prompting in Large Language Models — 书架上 1 本书引过它
- Conditional Prompt Learning for Vision-Language Models — 书架上 1 本书引过它
- DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines — 书架上 1 本书引过它
- EvoPrompt: Connecting LLMs with Evolutionary Algorithms Yields Powerful Prompt Optimizers — 书架上 1 本书引过它
- Generated Knowledge Prompting for Commonsense Reasoning — 书架上 1 本书引过它
- Interaction Scaling: Grounding the Third Axis of Test-Time Compute — 书架上 1 本书引过它
- Large Language Models Are Human-Level Prompt Engineers — 书架上 1 本书引过它
- Large Language Models as Optimizers — 书架上 1 本书引过它
- Learning to Prompt for Vision-Language Models — 书架上 1 本书引过它
- MCP-Zero: Active Tool Discovery for Autonomous LLM Agents — 书架上 1 本书引过它
- MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework — 书架上 1 本书引过它
- Not what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection — 书架上 1 本书引过它
- Prefix-Tuning: Optimizing Continuous Prompts for Generation — 书架上 1 本书引过它
- Rethinking the Role of Demonstrations: What Makes In-Context Learning Work? — 书架上 1 本书引过它
- RLPrompt: Optimizing Discrete Text Prompts with Reinforcement Learning — 书架上 1 本书引过它
- Segment Anything — 书架上 1 本书引过它
- Show Your Work: Scratchpads for Intermediate Computation with Language Models — 书架上 1 本书引过它
- TabLLM: Few-shot Classification of Tabular Data with Large Language Models — 书架上 1 本书引过它
- Understanding R1-Zero-Like Training: A Critical Perspective — 书架上 1 本书引过它
- When "A Helpful Assistant" Is Not Really Helpful: Personas in System Prompts Do Not Improve Performances of Large Language Models — 书架上 1 本书引过它
eval —— 基准、评测方法、评判模型、可观测性(27 篇)
- Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks — 书架上 9 本书引过它
- QLoRA: Efficient Finetuning of Quantized LLMs — 书架上 4 本书引过它
- Are Emergent Abilities of Large Language Models a Mirage? — 书架上 3 本书引过它
- Measuring Massive Multitask Language Understanding — 书架上 3 本书引过它
- Deep Residual Learning for Image Recognition — 书架上 2 本书引过它
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models — 书架上 2 本书引过它
- MTEB: Massive Text Embedding Benchmark — 书架上 2 本书引过它
- Ragas: Automated Evaluation of Retrieval Augmented Generation — 书架上 2 本书引过它
- Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference — 书架上 1 本书引过它
- Evaluating Very Long-Term Conversational Memory of LLM Agents — 书架上 1 本书引过它
- Gaussian Error Linear Units (GELUs) — 书架上 1 本书引过它
- Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context — 书架上 1 本书引过它
- GPQA: A Graduate-Level Google-Proof Q&A Benchmark — 书架上 1 本书引过它
- How Much Do LLMs Hallucinate in Document Q&A Scenarios? A 172-Billion-Token Study Across Temperatures, Context Lengths, and Hardware Platforms — 书架上 1 本书引过它
- Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena — 书架上 1 本书引过它
- Large Language Models as Optimizers — 书架上 1 本书引过它
- Llama 2: Open Foundation and Fine-Tuned Chat Models — 书架上 1 本书引过它
- LLaMA: Open and Efficient Foundation Language Models — 书架上 1 本书引过它
- On the Opportunities and Risks of Foundation Models — 书架上 1 本书引过它
- Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone — 书架上 1 本书引过它
- Robust Speech Recognition via Large-Scale Weak Supervision — 书架上 1 本书引过它
- Segment Anything — 书架上 1 本书引过它
- TabLLM: Few-shot Classification of Tabular Data with Large Language Models — 书架上 1 本书引过它
- The Leaderboard Illusion — 书架上 1 本书引过它
- Training Verifiers to Solve Math Word Problems — 书架上 1 本书引过它
- What Matters in Transformers? Not All Attention is Needed — 书架上 1 本书引过它
- When "A Helpful Assistant" Is Not Really Helpful: Personas in System Prompts Do Not Improve Performances of Large Language Models — 书架上 1 本书引过它
transformer —— 注意力、位置编码、架构本体(26 篇)
- Attention Is All You Need — 书架上 10 本书引过它
- LoRA: Low-Rank Adaptation of Large Language Models — 书架上 7 本书引过它
- Training Compute-Optimal Large Language Models — 书架上 7 本书引过它
- An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale — 书架上 3 本书引过它
- RoFormer: Enhanced Transformer with Rotary Position Embedding — 书架上 3 本书引过它
- BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding — 书架上 2 本书引过它
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness — 书架上 2 本书引过它
- BERTopic: Neural topic modeling with a class-based TF-IDF procedure — 书架上 1 本书引过它
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models — 书架上 1 本书引过它
- Decision Transformer: Reinforcement Learning via Sequence Modeling — 书架上 1 本书引过它
- DeepSeek-V3 Technical Report — 书架上 1 本书引过它
- DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models — 书架上 1 本书引过它
- Efficient Few-Shot Learning Without Prompts — 书架上 1 本书引过它
- Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer — 书架上 1 本书引过它
- Fast Transformer Decoding: One Write-Head is All You Need — 书架上 1 本书引过它
- GLM: General Language Model Pretraining with Autoregressive Blank Infilling — 书架上 1 本书引过它
- GLU Variants Improve Transformer — 书架上 1 本书引过它
- GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints — 书架上 1 本书引过它
- Mamba: Linear-Time Sequence Modeling with Selective State Spaces — 书架上 1 本书引过它
- MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention — 书架上 1 本书引过它
- On Layer Normalization in the Transformer Architecture — 书架上 1 本书引过它
- Show Your Work: Scratchpads for Intermediate Computation with Language Models — 书架上 1 本书引过它
- Textbooks Are All You Need — 书架上 1 本书引过它
- Training Verifiers to Solve Math Word Problems — 书架上 1 本书引过它
- Unsupervised Representation Learning with Deep Convolutional Generative Adversarial Networks — 书架上 1 本书引过它
- What Matters in Transformers? Not All Attention is Needed — 书架上 1 本书引过它
theory —— 表示学习、涌现、可解释性、理论分析(26 篇)
- Are Emergent Abilities of Large Language Models a Mirage? — 书架上 3 本书引过它
- Efficient Estimation of Word Representations in Vector Space — 书架上 3 本书引过它
- Auto-Encoding Variational Bayes — 书架上 2 本书引过它
- BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding — 书架上 2 本书引过它
- Learning Transferable Visual Models From Natural Language Supervision — 书架上 2 本书引过它
- Tree of Thoughts: Deliberate Problem Solving with Large Language Models — 书 架上 2 本书引过它
- ALBERT: A Lite BERT for Self-supervised Learning of Language Representations — 书架上 1 本书引过它
- Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift — 书架上 1 本书引过它
- BERTopic: Neural topic modeling with a class-based TF-IDF procedure — 书架上 1 本书引过它
- Denoising Diffusion Probabilistic Models — 书架上 1 本书引过它
- Emergent Abilities of Large Language Models — 书架上 1 本书引过它
- GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints — 书架上 1 本书引过它
- Intrinsic Dimensionality Explains the Effectiveness of Language Model Fine-Tuning — 书架上 1 本书引过它
- Late Chunking: Contextual Chunk Embeddings Using Long-Context Embedding Models — 书架上 1 本书引过它
- Layer Normalization — 书架上 1 本书引过它
- Making Large Language Models Efficient Dense Retrievers — 书架上 1 本书引过它
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view Reconstruction — 书架上 1 本书引过它
- On Layer Normalization in the Transformer Architecture — 书架上 1 本书引过它
- Physics-based Deep Learning — 书架上 1 本书引过它
- Reinforcement Learning for Generative AI: State of the Art, Opportunities and Open Research Challenges — 书架上 1 本书引过它
- RLPrompt: Optimizing Discrete Text Prompts with Reinforcement Learning — 书架上 1 本书引过它
- Robust Speech Recognition via Large-Scale Weak Supervision — 书架上 1 本书引过它
- Root Mean Square Layer Normalization — 书架上 1 本书引过它
- SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training — 书架上 1 本书引过它
- Textbooks Are All You Need — 书架上 1 本书引过它
- Will we run out of data? Limits of LLM scaling based on human-generated data — 书架上 1 本书引过它
post-training —— 指令微调、RLHF、DPO 一族、推理模型(19 篇)
- Language Models are Few-Shot Learners — 书架上 6 本书引过它
- Training language models to follow instructions with human feedback — 书架上 6 本书引过它
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning — 书架上 5 本书引过它
- Direct Preference Optimization: Your Language Model is Secretly a Reward Model — 书架上 3 本书引过它
- Precise Zero-Shot Dense Retrieval without Relevance Labels — 书架上 3 本书引过它
- Proximal Policy Optimization Algorithms — 书架上 2 本书引过它
- Agent-Computer Observation Interfaces Enable Dynamic Computer Use — 书架上 1 本书引过它
- Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference — 书架上 1 本书引过它
- DAPO: An Open-Source LLM Reinforcement Learning System at Scale — 书架上 1 本书引过它
- GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization — 书架上 1 本书引过它
- Group Sequence Policy Optimization — 书架上 1 本书引过它
- Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena — 书架上 1 本书引过它
- Large Language Models Are Human-Level Prompt Engineers — 书架上 1 本书引过它
- Let's Verify Step by Step — 书架上 1 本书引过它
- MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention — 书架上 1 本书引过它
- Not what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection — 书架上 1 本书引过它
- ORPO: Monolithic Preference Optimization without Reference Model — 书架上 1 本书引过它
- Rethinking Inference-Time Scaling: Efficiency Limits and Linguistic Signals — 书架上 1 本书引过它
- Understanding R1-Zero-Like Training: A Critical Perspective — 书架上 1 本书引过它
multimodal —— 视觉语言、跨模态、图文对齐(17 篇)
- An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale — 书架上 3 本书引过它
- Deep Residual Learning for Image Recognition — 书架上 2 本书引过它
- Learning Transferable Visual Models From Natural Language Supervision — 书架上 2 本书引过它
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models — 书架上 1 本书引过它
- Conditional Prompt Learning for Vision-Language Models — 书架上 1 本书引过它
- DAPO: An Open-Source LLM Reinforcement Learning System at Scale — 书架上 1 本书引过它
- Denoising Diffusion Probabilistic Models — 书架上 1 本书引过它
- Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context — 书架上 1 本书引过它
- Learning to Prompt for Vision-Language Models — 书架上 1 本书引过它
- Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model — 书架上 1 本书引过它
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view Reconstruction — 书架上 1 本书引过它
- On the Importance of Noise Scheduling for Diffusion Models — 书架上 1 本书引过它
- Segment Anything — 书架上 1 本书引过它
- Segment Anything Model for Medical Images? — 书架上 1 本书引过它
- SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training — 书架上 1 本书引过它
- The Curse of Recursion: Training on Generated Data Makes Models Forget — 书架上 1 本书引过它
- Unsupervised Representation Learning with Deep Convolutional Generative Adversarial Networks — 书架上 1 本书引过它
rag —— 检索增强、向量检索、重排、GraphRAG(13 篇)
- Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks — 书架上 9 本书引过它
- Lost in the Middle: How Language Models Use Long Contexts — 书架上 4 本书引过它
- From Local to Global: A Graph RAG Approach to Query-Focused Summarization — 书架上 3 本书引过它
- Precise Zero-Shot Dense Retrieval without Relevance Labels — 书架上 3 本书引过它
- Dense X Retrieval: What Retrieval Granularity Should We Use? — 书架上 2 本书引过它
- MTEB: Massive Text Embedding Benchmark — 书架上 2 本书引过它
- Ragas: Automated Evaluation of Retrieval Augmented Generation — 书架上 2 本书引过它
- Evaluating Very Long-Term Conversational Memory of LLM Agents — 书架上 1 本书引过它
- Late Chunking: Contextual Chunk Embeddings Using Long-Context Embedding Models — 书架上 1 本书引过它
- Making Large Language Models Efficient Dense Retrievers — 书架上 1 本书引过它
- MCP-Zero: Active Tool Discovery for Autonomous LLM Agents — 书架上 1 本书引过它
- Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge — 书架上 1 本书引过它
- User as Code: Executable Memory for Personalized Agents — 书架上 1 本书引过它
agents —— 工具调用、规划、多智能体、记忆(12 篇)
- ReAct: Synergizing Reasoning and Acting in Language Models — 书架上 6 本书引过它
- Agent-Computer Observation Interfaces Enable Dynamic Computer Use — 书架上 1 本书引过它
- DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models — 书架上 1 本书引过它
- DeepSeekMath-V2: Towards Self-Verifiable Mathematical Reasoning — 书架上 1 本书引过它
- Evaluating Very Long-Term Conversational Memory of LLM Agents — 书架上 1 本书引过它
- Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model — 书架上 1 本书引过它
- MCP-Zero: Active Tool Discovery for Autonomous LLM Agents — 书架上 1 本书引过它
- MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework — 书架上 1 本书引过它
- RLVP: Penalize the Path, Reward the Outcome — 书架上 1 本书引过它
- The Latent Bridge: A Continuous Slow-Fast Channel for Real-Time Game Agents — 书架上 1 本书引过它
- User as Code: Executable Memory for Personalized Agents — 书架上 1 本书引过它
- Whose Side Is Your Agent On? Multi-Party Principal Loyalty in LLM Agents — 书架上 1 本书引过它
systems —— 训练系统、并行、分布式、硬件(11 篇)
- Attention Is All You Need — 书架上 10 本书引过它
- Large Language Models are Zero-Shot Reasoners — 书架上 4 本书引过它
- Precise Zero-Shot Dense Retrieval without Relevance Labels — 书架上 3 本书引过它
- Einstein's Patents and Inventions — 书架上 1 本书引过它
- Fast Transformer Decoding: One Write-Head is All You Need — 书架上 1 本书引过它
- Large Language Models Cannot Self-Correct Reasoning Yet — 书架上 1 本书引过它
- Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model — 书架上 1 本书引过它
- On a new statistical technique for the real-time recognition of ultra-low multiplicity astrophysical neutrino burst — 书架上 1 本书引过它
- Robust Speech Recognition via Large-Scale Weak Supervision — 书架上 1 本书引过它
- TabLLM: Few-shot Classification of Tabular Data with Large Language Models — 书架上 1 本书引过它
- ZeRO: Memory Optimizations Toward Training Trillion Parameter Models — 书架上 1 本书引过它
efficiency —— 量化、蒸馏、剪枝、推理加速、长上下文(9 篇)
- Lost in the Middle: How Language Models Use Long Contexts — 书架上 4 本书引过它
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness — 书架上 2 本书引过它
- DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models — 书架上 1 本书引过它
- Distilling the Knowledge in a Neural Network — 书架上 1 本书引过它
- FineScope : SAE-guided Data Selection Enables Domain Specific LLM Pruning and Finetuning — 书架上 1 本书引过它
- Late Chunking: Contextual Chunk Embeddings Using Long-Context Embedding Models — 书架上 1 本书引过它
- MiniLLM: On-Policy Distillation of Large Language Models — 书架上 1 本书引过它
- Models Take Notes at Prefill: KV Cache Can Be Editable and Composable — 书架上 1 本书引过它
- The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural Networks — 书架上 1 本书引过它
safety —— 对齐、越狱、隐私、幻觉与事实性(8 篇)
- Training language models to follow instructions with human feedback — 书架上 6 本书引过它
- Ragas: Automated Evaluation of Retrieval Augmented Generation — 书架上 2 本书引过它
- How Much Do LLMs Hallucinate in Document Q&A Scenarios? A 172-Billion-Token Study Across Temperatures, Context Lengths, and Hardware Platforms — 书架上 1 本书引过它
- Llama 2: Open Foundation and Fine-Tuned Chat Models — 书架上 1 本书引过它
- MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework — 书架上 1 本书引过它
- Not what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection — 书架上 1 本书引过它
- Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone — 书架上 1 本书引过它
- Unsupervised Representation Learning with Deep Convolutional Generative Adversarial Networks — 书架上 1 本书引过它