Software Engineer
Babel Street
| Company | Babel Street |
| Category | Engineering |
| Location | Somerville |
| Remote | On-site (inferred) |
| Employment | Not stated |
| Level | Not stated |
| Salary | Not stated by the employer |
| Posted | 20 Jul 2026 |
| Last verified | 30 Jul 2026 |
| Source | Employer career page (greenhouse) |
Description
Babel Street is the trusted technology partner for the world’s most advanced identity intelligence and risk operations. We deliver advanced AI and data analytics solutions providing unmatched, analysis-ready data regardless of language, proactive risk identification, 360-degree insights, high-speed automation, and seamless integration into existing systems. Babel Street empower s government and commercial organizations to transform high-stakes identity and risk operations into a strategic advantage . The actionable insights we deliver safeguard lives and protect critical assets around the world . Babel Street is headquartered in Reston, Virginia , with regional offices in Boston, MA and Cleveland, OH, and international offices in Australia, Canada, Israel, Japan, and the U.K. For more information, visit www. babelstreet.com . ROLE SUMMARY:
As an early-career Engineer on the Image & Computer Vision AI team, you will support the development and deployment of computer vision capabilities that power Babel Street’s intelligence applications. You will help build systems that extract, analyze, and reason over visual data, including image search, object and scene understanding, facial matching workflows, geolocation support, and multimodal intelligence features.
This role is well suited for someone with foundational experience in computer vision, image processing, machine learning, or applied AI who is ready to grow in a hands-on engineering environment. You will work with senior engineers and cross-functional partners to implement, test, and improve reliable vision capabilities, including integration with multimodal LLM systems that allow users to search and reason over images using natural language.
This is a hybrid role to be based out of either our Reston, VA/Washington DC office or our Somerville MA office.
ROLE FOCUS;
This role spans three practical execution areas:
Computer Vision & Image Analytics
You will help implement and maintain image analytics pipelines that support facial matching, object detection, scene understanding, and image similarity. This includes supporting image preprocessing, feature extraction, model inference, evaluation, and performance improvements under the guidance of more senior team members.
Geospatial & Location Inference from Imagery
You will assist with capabilities that infer location, context, or environmental attributes from imagery by using visual cues, metadata, and learned representations. This may include supporting image-based geolocation, landmark recognition, and contextual scene analysis used in intelligence workflows.
Multi-Modal AI & Image Search
You will contribute to multimodal AI systems that combine vision models with LLMs, embeddings, and retrieval pipelines to support natural-language search and reasoning over images and image collections. You will help integrate visual understanding into broader intelligence applications and workflows, including supporting entity and event extraction from image-based intelligence data.
KEY RESPONSIBILITIES:
Assist in building and maintaining computer vision pipelines for image ingestion, preprocessing, inference, and evaluation.
Support facial matching and identity-related vision workflows in accordance with accuracy, safety, and compliance requirements.
Help develop, test, and improve object detection, image similarity, and scene understanding models.
Contribute to image-based geolocation and location i