About the Role
Join a small, focused team of researchers and engineers working at the frontier of learning and memory. As a Research Scientist, you’ll design experiments, develop new recipes, build evals, and shape the product used by some of the world's leading tech and AI companies.
Responsibilities
- Design experiments, develop new recipes, and build evals.
- Shape the product used by some of the world's leading tech and AI companies.
- Design and evaluate methods for encoding large, heterogeneous document corpora into compact parametric memory.
- Develop self-study pipelines that allow models to reflect on and consolidate new context.
- Tackle catastrophic forgetting, sequential updates, knowledge conflicts, and tradeoffs between in-weights memory and agentic retrieval.
- Explore reinforcement learning methods that let models improve from interaction and feedback in real deployment settings.
- Empirically study how model capacity, data scale, and compute interact.
- Develop scaling laws that inform our product roadmap.
Requirements
- A deep background in machine learning, with strong fundamentals in inference serving systems, KV cache design, or latency-sensitive model deployment.
- A track record of rigorous ML research (publications, strong open-source contributions, or equivalent demonstrated depth).
- Extensive experience in at least one area directly relevant to our work: continual learning, memory architectures, test-time training (TTT), parameter-efficient finetuning, context compression, retrieval, synthetic data, distillation, or agents.
- Comfort working up and down the stack, understanding both the research question and the system it runs on.
- Strong technical communication skills, with the ability to explain complex ideas simply and engage in high-bandwidth, generative technical conversation.
- Experience bridging research and product, shipping things that real users interact with.
- Familiarity with LLM training infrastructure.
Skills
- Machine learning
- Inference serving systems
- KV cache design
- Latency-sensitive model deployment
- ML research
- Continual learning
- Memory architectures
- Test-time training (TTT)
- Parameter-efficient finetuning
- Context compression
- Retrieval
- Synthetic data
- Distillation
- Agents
- Technical communication
- LLM training infrastructure
Location
- San Francisco
Work Type
- In-person
Experience Level
- Research Scientist
Salary/Compensations
- Competitive cash compensation
- Startup equity
About the Company
- Today’s AI is a brilliant stranger: it can solve the world’s hardest math problems, but it knows next to nothing about you and your work. It rereads your files to answer even basic questions, burns an enormous amount of tokens when sifting through large corpuses, and between sessions, it retains scraps at best.
- We train models to study your world and anticipate your questions in advance, forming engrams: compact memories that capture your knowledge and history. Our approach opens a new axis of scaling. The more we study your context at training time, the better we become at inference time.
- We're already working with leaders in AI like Microsoft, Notion, and Harvey, and just raised $98M from General Catalyst, Kleiner Perkins, Sequoia, Factory, Modern, Amplify, Neo and others. Our investors and advisors include Assaf Rappaport, Andrej Karpathy, and Pieter Abbeel.
- AI has spent years learning everything about the world. Now it should learn something about yours.
- Our team is passionate about the problems we are solving. We work up and down the stack, and the line between research and engineering is blurry by design. We're pragmatic and problem-driven, collaborative to our core, and hold a high bar for everything we ship.
Equal Opportunity
- Engram is an equal opportunity employer. We’re building a team that reflects a range of backgrounds and perspectives, and we welcome applicants regardless of race, color, religion, national origin, gender, gender identity, sexual orientation, age, disability, or veteran status.
