I'm interested in scalable software.
Current intern at IBM working on agents and distributed LLM inference.
Recently I've been contributing to Opal, an open source simulator for exploring distributed LLM inference policies (for systems like llm-d and Dynamo) without consuming GPU or storage infrastructure resources.
Always happy to chat - reach me at james.yan2028@gmail.com.