[1]Vovk, Gammerman & Shafer (2005). Algorithmic Learning in a Random World. Springer.
[2]Clopper & Pearson (1934). The Use of Confidence or Fiducial Limits Illustrated in the Case of the Binomial. Biometrika.
[3]Lei, G'Sell, Rinaldo, Tibshirani & Wasserman (2018). Distribution-Free Predictive Inference for Regression. JASA.
[4]Angelopoulos, Bates, Candès, Jordan & Lei (2021). Learn then Test: Calibrating Predictive Algorithms to Achieve Risk Control. arXiv:2110.01052.
[5]Bates, Angelopoulos, Lei, Malik & Jordan (2021). Distribution-Free, Risk-Controlling Prediction Sets. JACM.
[6]Min et al. (2023). FActScore: Fine-grained Atomic Evaluation of Factual Precision. EMNLP.
[7]Park, O'Brien, Cai, Morris, Liang & Bernstein (2023). Generative Agents: Interactive Simulacra of Human Behavior. UIST.
[8]Packer et al. (2023). MemGPT: Towards LLMs as Operating Systems. arXiv:2310.08560.
[9]Wu, Wang, Yu, Zhang, Chang & Yu (2024). LongMemEval: Benchmarking Chat Assistants on Long-Term Interactive Memory. arXiv:2410.10813.
[10]Maharana et al. (2024). Evaluating Very Long-Term Conversational Memory of LLM Agents. ACL.
[11]Chhikara, Khant, Aryan, Singh & Yadav (2025). Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory. arXiv:2504.19413.
[12]Pan et al. (2024). LLMLingua-2: Data Distillation for Efficient and Faithful Task-Agnostic Prompt Compression. ACL Findings.
[13]Xu, Liang, Mei, Gao, Tan & Zhang (2025). A-MEM: Agentic Memory for LLM Agents. arXiv:2502.12110.
[14]Tacheny (2025). Geometric Dynamics of Agentic Loops in Large Language Models. arXiv preprint.
[15]Angelopoulos, Bates, Fisch, Lei & Schuster (2024). Conformal Risk Control. ICLR.
[16]Ramdas, Grünwald, Vovk & Shafer (2023). Game-Theoretic Statistics and Safe Anytime-Valid Inference. Statistical Science.