About the company
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs, delivering industry-leading training and inference speeds. Cerebras works with leading model labs, global enterprises, and AI-native startups.
Responsibilities
- Develop x86 & ARM software to expose next-generation hardware IO capabilities for AI/HPC application teams
- Govern a generic IO API with multiple internal users
- Develop control and configuration subsystems directly interacting with Cerebras hardware
- Drive network performance debug of large AI clusters
- Gather and analyze network statistics and packet traces to root cause bottlenecks
- Develop tools/telemetry for increasing visibility into the network and IO datapath
- Optimize CPU/mem utilization leveraging kernel bypass and zero-copy techniques
- Integrate leading edge networking technologies and protocols
- Lead cross-functional technical projects
Requirements
- Master's/PhD in Computer Science or Electrical Engineering + 1 year industry experience, OR 3+ years industry experience
- Experience in large software environments
- Embedded systems, HW/SW co-design, and some driver development
- Network protocol familiarity (TCP, RoCE) and network debug tools such as Wireshark
- Network switch environment familiarity (Arista, Juniper, etc.)
- Detail-oriented and eager to learn
Conditions
- Build a breakthrough AI platform beyond the constraints of the GPU
- Publish and open source cutting-edge AI research
- Work on one of the fastest AI supercomputers in the world
- Job stability with startup vitality
- Non-corporate work culture