Skip to main content
Tallo logoTallo logo

Find Jobs

Find Jobs Near You – Available Work in Your Location

Skip to job details

Back to Results

Apply for this opportunity

To apply for this job, you'll continue to an external website or email application.

Interval Partners, LP

Data Scientist / Data Engineer, Alternative Data and AI

Choose a Location

This role is available in multiple locations. Pick one to apply.

Career Insights for Natural Language Processing Engineer

See where this job fits in the broader career landscape. Knowing your career path helps you see what's possible from here.

Scorecard

Based on New Jersey data

Review key factors to help you decide if this role fits your goals. How is this calculated?

Were these scores useful?

What they do

A Natural Language Processing Engineer specializes in developing and implementing algorithms and models tailored for understanding, processing, and generating natural language text. They utilize methodologies such as tokenization, parsing, named entity recognition, part-of-speech tagging, and other NLP techniques to perform tasks including text classification, chatbot development, and other applications where the primary input or output is natural language text.

$114,982 / year median in New Jersey

Explore Career

Job Description

Data Scientist / Data Engineer, Alternative Data and AI at Interval Partners, LP Data Scientist / Data Engineer, Alternative Data and AI at Interval Partners, LP in Jersey City, New Jersey Posted in about 24 hours ago.

Type:

full-time

INTERVAL PARTNERS

Data Scientist / Data Engineer, Alternative Data and AI 1-3 Years Experience ???? New York, NY, Midtown East, Onsite ? Full-Time ?? jobs@intervalpartners.com

ABOUT THE ROLE

Interval Partners is a multi-billion-dollar alternative investment firm located in Midtown Manhattan. We are seeking a Data Scientist/Engineer to join our team. This role will report to the firm's Data Engineer/Developer and President. This is a hands-on, ownership-oriented role focused first on data engineering: building reliable pipelines, maintaining curated datasets, and making alternative data ready for analyst and portfolio manager use. The role also requires strong judgment about what each dataset measures, where its limitations are, and how LLMs, AI agents, and tool-based workflows can make analysts faster. You will work directly with portfolio managers and senior analysts on questions tied to live investment decisions, with reliable data and targeted AI solutions at the center of the work.

KEY RESPONSIBILITIES

Alternative Data Acquisition, Ingestion & Pipeline Maintenance

  • Build, maintain, and improve Python-based pipelines for ingesting, processing, and visualizing structured datasets, including alternative data such as credit/debit card transactions, point-of-sale data, and internal data assets.
  • Collaborate on ingestion workflows that are reliable, repeatable, observable, and easy to maintain, with clear treatment of schema changes, late-arriving data, vendor restatements, duplicate records, and missing values.
  • Develop automated checks for data completeness, consistency, timeliness, and accuracy, and create clear escalation paths when pipeline failures or data quality issues occur. Data Cleaning, Transformation & Readiness
  • Transform raw data into clean, well-structured, analysis-ready outputs that map to relevant business metrics, key performance indicators, and analyst research workflows.
  • Perform data profiling, normalization, enrichment, deduplication, entity resolution, outlier handling, and quality remediation to improve downstream usability.
  • Understand the meaning, lineage, limitations, and caveats of each dataset, and clearly document assumptions, coverage gaps, definitions, and known quality constraints. Analyst Enablement & Data Understanding
  • Work closely with analysts and portfolio managers to understand research questions, translate them into data requirements, and deliver well-documented datasets, extracts, and analyses.
  • Dig into the drivers behind trends observed in the data, helping analysts distinguish durable signals from noise, one-off effects, data artifacts, or coverage changes.
  • Surface data-driven alerts, explainable anomalies, and relevant changes in key metrics where the underlying data quality and business interpretation are well understood.
  • Communicate technical findings, data caveats, statistical context, and limitations clearly to both technical and non-technical stakeholders. AI Readiness & Practical AI Use Cases
  • Maintain data assets in formats that can be safely and effectively used by analytics tools, LLM applications, AI agents, and retrieval or tool-based workflows.
  • Demonstrate a good conceptual understanding of large language models, AI agents, tool use, retrieval-augmented generation, embeddings, structured outputs, and prompt-driven workflows.
  • Identify practical AI-enabled use cases that improve analyst efficiency, such as conversational data exploration, automated research summaries, data quality explanations, metric lookup, and hypothesis triage.
  • Partner with technology teams to ensure that AI solutions are grounded in clean, documented, well-permissioned, and trustworthy data rather than treating AI development as the primary responsibility of the role. Analytical Methods & Model Evaluation
  • Apply appropriate statistical and time-series techniques to support KPI forecasting, anomaly detection, trend analysis, and signal evaluation when required by analyst use cases.
  • Conduct disciplined backtesting and validation of datasets, signals, and model outputs, with attention to overfitting, data revisions, and signal stability.
  • Document model assumptions, evaluation results, confidence ranges, and limitations in a way that supports informed analyst decision-making.
QUALIFICATIONS

Required

  • Hands-on experience with Python for data engineering and analysis, including pandas, numpy, and data validation techniques.
  • Bachelor's degree in Computer Science, Data Engineering, Statistics, Applied Mathematics, Engineering, or a related quantitative discipline.
  • Ability to understand business context, map raw data observations to meaningful metrics, and explain data limitations clearly.
  • Familiarity with LLM tooling and interest in applying it to analyst workflows.
  • Solid grounding in statistics, time-series analysis, forecasting, anomaly detection, and disciplined backtesting practices.
  • Excellent written and verbal communication skills, with the ability to present data issues, assumptions, and technical findings clearly to analysts and decision-makers.
  • Highly organized, self-directed, and comfortable maintaining multiple datasets and pipelines in a fast-paced environment. Preferred
  • Experience working with alternative data vendors, investment research datasets, financial datasets, or other high-volume third-party data sources.
  • Proven experience building, operating, and maintaining end-to-end data pipelines in a production or business-critical environment.
  • Experience preparing datasets for LLM, RAG, agentic analytics, semantic search, or conversational data exploration use cases.

This is a fully onsite 5 days a week role based in our Midtown office.

Benefits:

Full medical & vision, 401(k). Compensation range is $85,000-$100,000.

Benefits

  • 401(k) Plans
  • Dental Insurance