
AI Researcher
Remote
Full Time
#Engineering
#Python
#PyTorch
#Machine Learning
#Optimization
Autify, Inc. is a San Francisco-based startup with a mission to empower human creativity through advanced technology. Since our founding by the first Japanese team to graduate from the Alchemist Accelerator, we have been dedicated to transforming software quality assurance. We launched our AI-powered no-code test automation platform in 2019 and have since grown our team to over 100 professionals. With the recent release of our GenAI agent, Genesis, we are doubling down on our commitment to intelligent automation. As we expand our global reach, we are looking for a Senior AI Researcher to join our international engineering team and help us push the boundaries of what our platform can achieve.
Key outcomes
- Research, prototype, and implement enhancements for our AI agent systems, with a focus on adaptive behavior and contextual understanding.
- Design and train custom models, including small language models and hybrid architectures, for local execution.
- Optimize memory usage and performance for AI features deployed in desktop or edge environments.
- Collaborate with cross-functional teams to identify and develop new use cases for intelligent agents.
- Evaluate and integrate the latest advancements in multi-agent systems, LLMs, and retrieval-augmented generation techniques.
- Stay at the forefront of AI developments to provide insights that influence our product roadmap.
- Develop internal tools to support model deployment and ongoing experimentation.
Requirements
- At least 5 years of experience in applied machine learning or AI research, preferably within product-focused environments.
- Deep technical knowledge of modern transformers, LLMs, and techniques such as RAG, instruction tuning, or contextual embedding.
- Proven experience in distilling or customizing large models using methods like quantization, pruning, or LoRA.
- Strong proficiency in Python and ML frameworks such as PyTorch or Hugging Face Transformers.
- A solid engineering foundation with the ability to build production-ready prototypes and deploy ML systems.
- Familiarity with multi-modal agent systems, edge ML deployments, or local inference techniques.
- Clear and effective communication skills in English.
Preferred qualifications
- Experience working with embedded or small language models like LLaMA variants or TinyML.
- Background in autonomous agents, knowledge graphs, or few-shot learning.
- Familiarity with privacy-preserving AI techniques such as differential privacy or federated learning.
- Experience with desktop-based or real-time inference environments using tools like TensorRT or ONNX.
- A history of contributions to open-source AI projects.
Compensation
We offer a full-time remote work environment. Our benefits include the flexibility of working remotely.
How to apply
If you are a product-minded researcher who enjoys solving complex problems with a pragmatic approach, we would love to hear from you. Please submit your application to join our team as we continue to innovate in the software quality assurance space.







