Position: Data Scientist
Location: Irving, TX
Experience: 4–7 years of relevant experience
Candidate Requirements
- Local candidates only or candidates located within a 100-mile radius of Irving, TX.
- Genuine consultants only with strong communication skills.
- Must have a valid LinkedIn profile that is consistent with the candidate's resume.
- Candidates must be able to demonstrate strong hands-on technical experience in the required skill set.
- Candidates open to relocation may also be considered.
Required Education
- Bachelor's degree in Computer Science, Information Technology, or a related field, or equivalent work experience.
Required Skills & Responsibilities
- Join the Autonomy and Automation – Data Pipeline Team supporting annotation tooling used for autonomous machine development.
- Develop, enhance, maintain, and support annotation tools used by developers and data annotators.
- Strong hands-on experience with Python for application and tool development.
- Experience with UI/UX development and designing user-friendly tool interfaces.
- Experience working with cloud-based environments; cloud experience is highly preferred.
- Experience with API development and integrations, including communication between multiple tools and systems.
- Strong GitHub experience, including source control and collaborative development workflows.
- Experience with CI/CD workflows to ensure code and tooling are properly tested before deployment.
- Experience with Docker/containerization is preferred.
- Experience with image processing is a plus, including the ability to understand and manipulate images within development tools.
- Background in Computer Vision and/or Machine Learning is highly preferred.
- Understanding of machine learning workflows and practical development processes.
- Familiarity with agentic coding practices is preferred.
- Ability to understand how end users interact with and utilize development tools.
Top Required Skills
- Python
- Agentic Coding Practices
- Machine Learning Workflows
- UI/UX
- GitHub
- API Development
- CI/CD
- Cloud Technologies
Role Breakdown
- Approximately 60% of the role will involve hands-on development, primarily using Python and UI/UX.
- Approximately 40% will involve meetings, collaboration, learning, and understanding user requirements and workflows.
Project Overview
The annotation tool is used to support Caterpillar's autonomous machine training initiatives. Tasks are loaded into the tool, where developers/annotators work with camera or image data representing real-world environments.
Users draw boundaries around objects within the images or camera views to identify and define objects such as trucks, people, and other objects in the environment. The annotated data is then used to support the training and development of autonomous machines.
The selected candidate will be responsible for developing and enhancing the tooling that enables these annotation workflows, while ensuring the application is reliable, user-friendly, and properly tested before deployment.
Interview Process
- Initial screening
- Technical interview
- Coding exercise