Performance Engineer
2 days ago
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs.
Cerebras' current customers include global corporations across multiple industries, national labs, and top-tier healthcare systems. In January, we announced a multi-year, multi-million-dollar partnership with Mayo Clinic, underscoring our commitment to transforming AI applications across various fields. In August, we launched Cerebras Inference, the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services.
About The RoleEngineers on the inference performance team operate at the intersection of hardware and software, driving end-to-end model inference speed and throughput. Their work spans low-level kernel performance debugging and optimization, system-level performance analysis, performance modeling and estimation, and the development of tooling for performance projection and diagnostics.
Responsibilities- Build performance models (kernel-level, end-to-end) to estimate the performance of state of the art and customer ML models.
- Optimize and debug our kernel micro code and compiler algorithms to elevate ML model inference speed, throughput and compute utilization on the Cerebras WSE.
- Debug and understand runtime performance on the system and cluster.
- Develop tools and infrastructure to help visualize performance data collected from the Wafer Scale Engine and our compute cluster.
- Bachelors / Masters / PhD in Electrical Engineering or Computer Science.
- Strong background in computer architecture.
- Exposure to and understanding of low-level deep learning / LLM math.
- Strong analytical and problem-solving mindset.
- 3+ years of experience in a relevant domain (Computer Architecture, CPU/GPU Performance, Kernel Optimization, HPC).
- Experience working on CPU/GPU simulators.
- Exposure to performance profiling and debug on any system pipeline.
- Comfort with C++ and Python.
People who are serious about software make their own hardware. At Cerebras we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we've reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:
- Build a breakthrough AI platform beyond the constraints of the GPU.
- Publish and open source their cutting-edge AI research.
- Work on one of the fastest AI supercomputers in the world.
- Enjoy job stability with startup vitality.
- Our simple, non-corporate work culture that respects individual beliefs.
Read our blog: Five Reasons to Join Cerebras in 2025.
Apply today and become part of the forefront of groundbreaking advancements in AICerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.
This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.
-
Performance Engineer
1 week ago
Toronto, Ontario, Canada Veeva Systems Full timeVeeva Systems is a mission-driven organization and pioneer in industry cloud, helping life sciences companies bring therapies to patients faster. As one of the fastest-growing SaaS companies in history, we surpassed $2B in revenue in our last fiscal year with extensive growth potential ahead.At the heart of Veeva are our values: Do the Right Thing, Customer...
-
Sr Performance Engineer
6 days ago
Toronto, Ontario, Canada Tangerine Full timeRequisition ID: 232611Tangerine is Canada's leading direct bank. We offer flexible and accessible banking options, innovative products, and award-winning Client service. The reason why Tangerine employees come to work each day is to help Canadians live better lives. We focus on making a difference in our communities, and that includes our own internal...
-
Performance Test Engineer
5 days ago
Toronto, Ontario, Canada Atlantis IT Group Full timePerformance engineer with hands on exp in Jmeter8 + years of hands on experience in JMeter The candidate must have experience in Resiliency testing The candidate must have experience in Disaster recovery testing The candidate must have hands on experience with databases and writing SQL queries, Java and JavaScript, Must have experience the Application...
-
Dynatrace Performance Engineer
2 days ago
Toronto, Ontario, Canada Viva Tech Solutions Full timeQualifications:Bachelor's degree in Computer Science, Engineering, or related field.5+ years of experience in application performance engineering or a related role.Strong proficiency with Dynatrace.Solid knowledge of performance testing tools (e.g., JMeter, LoadRunner).Understanding of distributed systems, microservices, and cloud environments.Experience...
-
Performance Engineering
1 week ago
Toronto, Ontario, Canada Tata Consultancy Services Full timeInclusion without Exception:Tata Consultancy Services (TCS) is an equal opportunity employer, and embraces diversity in race, nationality, ethnicity, gender, age, physical ability, neurodiversity, and sexual orientation, to create a workforce that reflects the societies we operate in. Our continued commitment to Culture and Diversity is reflected in our...
-
Performance Test Engineer
6 hours ago
Toronto, Ontario, Canada BURGEON IT SERVICES Full timeRole: Performance Test EngineerExperience: 10+ Years (6+ Years Relevant)Location: Tiverton, Ontario, Canada (WFO – Onsite only)Domain: Utilities – ElectricityPlease share the resume at Roles & ResponsibilitiesMinimum 10+ years of experience in Performance Testing with hands-on execution.At least 3+ years of strong working experience in Apache JMeter ...
-
Senior ML Performance Engineer
1 week ago
Toronto, Ontario, Canada Lemurian Labs Full time**About UsAt Lemurian Labs, we're on a mission to bring the power of AI to everyone—without leaving a massive environmental footprint. We care deeply about the impact AI has on our society and planet, and we're building a solid foundation for its future, ensuring AI grows sustainably and responsibly. Innovation should help the world, not harm it.We are...
-
Performance Test Engineer
2 weeks ago
Toronto, Ontario, Canada Kumaran Systems Full timeSeeking an experienced Performance Engineer QE to join our Technology team, specializing in Non-Functional Requirements (NFR) testing and performance optimization for US Corporate Banking platforms. This role focuses on developing robust NFR test strategies, collaborating with application teams, and ensuring optimal system performance through effective tool...
-
RAN Performance and Optimization Engineer
2 weeks ago
Toronto, Ontario, Canada HiringAgents Full timeJob title: RAN Performance and Optimization EngineerClient: Myticas ConsultingLocation: Toronto, Ontario, Canada - On-SiteContract type: ContractContract duration:Salary:About The RoleMyticas Consulting is seeking a RAN Performance and Optimization Engineer to support a leading telecommunications provider in Toronto. In this role, you will focus on LTE and...
-
DevOps Performance Engineer
2 days ago
Toronto, Ontario, Canada AstraNorth Full timeLocation: Toronto. Required Skills: Digital : DevOps~Performance Engineering Experience: 6-8 yearsMust HaveGarbage Collection Analysis Server/Host Capacity AnalysisUnderstanding of RESTful API designUnderstanding of DevOps and Agile MethodologyUnderstanding of Three-tier web architecture pattern and Client-Server modelStrong foundation in programming and...