OpenAI
Anthropic
and other commercial LLM providers
AI Sandbox (Portkey)
GitHub Copilot
Claude PU Enterprise (faculty sponsor, staff)
see: https://dais.princeton.edu/resources
Anything LLM
Ollama
LM Studio
MLX (Apple Silicon)
vLLM
Before you identify which models are best for a given project, consider your requirements for cost, reproducibility, confidentiality, and compute capacity.
Confidential Data
Confidential Data
Confidential Data
Agentic coding assistants like Claude Code or Codex add the ability to work with a folder of files on your computer.
I code
It codes
Spec.md - reusable reference file about the larger goals, guidance, and requirements of your project. Usually written with/by the LLM.
Skills.md - Reusable description of a capability or resource
MCP and WebMCP (model context protocol) - An external resource or tool. For example tools to search Princeton's library catalog
Ponytail - A package of software engineering best practices and opinions
Pre-loaded MCP and skills for scientific literature and evidence synthesis (primarily life sciences).
Designed for computational analysis using a local machine and high-performance computing
Designed for reproducibility. All "artifacts" have a history and logging.
Antaripa Saha, Hamel Husain, and Hugo Bowne-Anderson
While the harness turns capability into work, it also adds complexity. How do we know that the model chose the right tool? That the tool-call led to improvement? Where did it add new errors? Evaluation is the craft of making those processes explicit and managable.
From the very beginning, in your specifications, call for code tests and evaluation.