D. Sculley et al., 《Hidden Technical Debt in Machine Learning Systems》, NeurIPS, 2015.
A. Vaswani et al., 《Attention Is All You Need》, NeurIPS, 2017.
J. Xin et al., 《DeeBERT: Dynamic Early Exiting for Accelerating BERT Inference》, ACL, 2020, arXiv:2004.12993.
E. Ries, The Lean Startup, Crown Business, 2011.
J. Humble & D. Farley, Continuous Delivery, Addison-Wesley, 2010.
N. Forsgren, J. Humble, and G. Kim, Accelerate: The Science of Lean Software and DevOps, IT Revolution, 2018.
T. Mikolov et al., 《Efficient Estimation of Word Representations in Vector Space》, arXiv:1301.3781, 2013.
D. Blei et al., 《Latent Dirichlet Allocation》, JMLR, 2003.
M. Grootendorst, 《BERTopic: Neural Topic Modeling with a Class-Based TF-IDF Procedure》, EMNLP, 2022.
J. S. Park et al., 《Generative Agents: Interactive Simulacra of Human Behavior》, arXiv:2304.03442, 2023.
ISO/IEC/IEEE International Standard - Systems and software engineering -- Life cycle processes -- Requirements engineering (IEEE Std 29148-2018), 2018, DOI: 10.1109/IEEESTD.2018.8559686.
UML 2.5 Specification, OMG, 2015.
E. Evans, Domain-Driven Design, Addison-Wesley, 2003.
P. Isola et al., 《Image-to-Image Translation with Conditional GANs》, CVPR, 2017.
H. Nguyen et al., 《Text-to-UI Generation for Web Applications》, UIST, 2023.
J. Nielsen, Usability Engineering, Morgan Kaufmann, 1993.
W3C, Web Content Accessibility Guidelines (WCAG) 2.2, 2023.
K. Beck, Test-Driven Development: By Example, Addison-Wesley, 2002.
GitHub, 《Copilot Trust Center》, 2023.
L. Richardson & S. Ruby, RESTful Web Services, O'Reilly, 2007.
S. Tilkov & S. Vinoski, 《Node.js: Using JavaScript to Build High-Performance Network Programs》, IEEE Internet Computing, 2010.
D. Hardt, 《The OAuth 2.0 Authorization Framework》, RFC 6749, 2012.
PCI Security Standards Council, Payment Card Industry Data Security Standard v4.0, 2022.
P. Lewis et al., 《Retrieval-Augmented Generation for Knowledge-Intensive NLP》, NeurIPS, 2020.
J. Johnson et al., 《Billion-Scale Similarity Search with GPUs》, IEEE Trans. Big Data, 2019.
J. Devlin et al., 《BERT: Pre-training of Deep Bidirectional Transformers》, NAACL, 2019.
S. Robertson & H. Zaragoza, 《The Probabilistic Relevance Framework: BM25 and Beyond》, Foundations and Trends in IR, 2009.
P. Li et al., 《RAGAS: Automated Evaluation of Retrieval-Augmented Generation》, arXiv:2402.01761, 2024.
S. Yao et al., 《ReAct: Synergizing Reasoning and Acting in Language Models》, ICLR, 2023.
Q. Wu et al., 《AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation》, arXiv:2308.08155, 2023.
OpenAI, 《Function Calling and Tool Use》, 2023.
C. Yeh et al., 《ChartLLM: A Benchmark for Chart Reasoning》, EMNLP, 2023.
A. Zhang et al., Dive into Deep Learning, 2021.
T. Gebru et al., 《Datasheets for Datasets》, CACM, 2021.
S. Mouli et al., 《Data Deduplication for Large Language Models》, arXiv:2307.03195, 2023.
Y. Wang et al., 《Self-Instruct: Aligning Language Models with Self-Generated Instructions》, ACL, 2023.
A. Gu et al., 《Domain Adaptation of Language Models via Continued Pretraining》, ACL, 2021.
R. Sennrich et al., 《Neural Machine Translation of Rare Words with Subword Units》, ACL, 2016.
S. Narayanan et al., 《Efficient Large-Scale Language Model Training on GPU Clusters》, MLSys, 2021.
E. Hu et al., 《LoRA: Low-Rank Adaptation of Large Language Models》, ICLR, 2022.
L. Ouyang et al., 《Training Language Models to Follow Instructions with Human Feedback》, NeurIPS, 2022.
S. Rafailov et al., 《Direct Preference Optimization: Your Language Model is Secretly a Reward Model》, NeurIPS, 2023.
T. Schick et al., 《Toolformer: Language Models Can Teach Themselves to Use Tools》, arXiv:2302.04761, 2023.
D. Patterson et al., 《Carbon Emissions and Large Neural Network Training》, arXiv:2104.10350, 2021.
H. Kwon et al., 《vLLM: Fast and Cheap LLM Serving with PagedAttention》, SOSP, 2023.
S. Amershi et al., 《Guidelines for Human-AI Interaction》, CHI, 2019, https://www.microsoft.com/en-us/research/publication/guidelines-for-human-ai-interaction/, accessed 2025-12-18.
Google PAIR, People + AI Guidebook, https://pair.withgoogle.com/guidebook/, accessed 2025-12-18.
T. Torres, Continuous Discovery Habits, https://www.producttalk.org/continuous-discovery-habits/, accessed 2025-12-18.
S. Bradner, 《Key words for use in RFCs to Indicate Requirement Levels》, RFC 2119, 1997, https://www.rfc-editor.org/rfc/rfc2119, accessed 2025-12-18.
Nielsen Norman Group, The User Experience of Customer-Service Chat: 20 Guidelines, https://www.nngroup.com/articles/chat-ux/, accessed 2025-12-19.
Nielsen Norman Group, Prompt Controls in GenAI Chatbots: 4 Main Uses and Best Practices, https://www.nngroup.com/articles/prompt-controls-genai/, accessed 2025-12-19.
Nielsen Norman Group, Testing AI with Real Design Scenarios: Evaluation Methodology and Prompts, https://www.nngroup.com/articles/testing-ai-methodology/, accessed 2025-12-19.
Nielsen Norman Group, Good from Afar, But Far from Good: AI Prototyping in Real Design Contexts, https://www.nngroup.com/articles/ai-prototyping/, accessed 2025-12-19.
Nielsen Norman Group, Leverage AI for Mock Tables and Charts When Testing Prototypes, https://www.nngroup.com/articles/ai-data-prototype-testing/, accessed 2025-12-19.
J. Schulman et al., 《Proximal Policy Optimization Algorithms》, arXiv:1707.06347, 2017, https://arxiv.org/abs/1707.06347.
P. F. Christiano et al., 《Deep Reinforcement Learning from Human Preferences》, NeurIPS, 2017, https://arxiv.org/abs/1706.03741.
D. Ziegler et al., 《Fine-Tuning Language Models from Human Preferences》, arXiv:1909.08593, 2019, https://arxiv.org/abs/1909.08593.
N. Stiennon et al., 《Learning to Summarize from Human Feedback》, NeurIPS, 2020, https://arxiv.org/abs/2009.01325.
Y. Bai et al., 《Constitutional AI: Harmlessness from AI Feedback》, arXiv:2212.08073, 2022, https://arxiv.org/abs/2212.08073.
D. Amodei et al., 《Concrete Problems in AI Safety》, arXiv:1606.06565, 2016, https://arxiv.org/abs/1606.06565.