2

Remote Data Infrastructure Jobs in New York (NOW HIRING)

Sr. Data Analyst - Remote

Manhattan, NY ยท On-site +1

$94K - $119K/yr

Details: Sr. Data Analyst Duration: Full time / Direct Hire Location: 100% Remote but candidate ... Maintain Insider Intelligence data infrastructure. * Develop reporting dashboards in PowerBI or ...

AI Data Engineer

Secaucus, NJ ยท Remote

$100K - $150K/yr

AI Data Engineer - Remote Bright Vision Technologies is a technology consulting and software ... Stay current with AI data infrastructure research and emerging open-source tools. Required ...

Data Engineer

Parsippany, NJ ยท Remote

$105K - $151K/yr

Deep acuity in leveraging the latest in big data and cloud infrastructure methodologies ... Position is Parsippany, NJ preferred; remote considered The US base salary range for this full-time ...

Data Engineer

Parsippany, NJ ยท Remote

$105K - $151K/yr

Deep acuity in leveraging the latest in big data and cloud infrastructure methodologies ... Position is Parsippany, NJ preferred; remote considered The US base salary range for this full-time ...

Data Analyst III

New York, NY ยท On-site +1

$114K - $142K/yr

Data at Brex The Data organization develops insights, models, and data infrastructure for teams ... As a perk, we also have up to four weeks per year of fully remote work! Responsibilities * Own the ...

Lead Data Engineer

New York, NY ยท Remote

$200K - $250K/yr

Tech Lead - Data Platform | Remote A fast-growing financial technology company is seeking a Tech Lead, Data Platform to design and build the core data infrastructure that powers advanced financial ...

Remote Role Responsibilities * Use frontier AI coding agents to complete and evaluate complex data ... Ability to evaluate model-generated data infrastructure and pipeline implementations. Preferred

Data Engineer II

Manhattan, NY ยท Remote

$117K - $140K/yr

Data Engineer II Location-Type: Remote (EST hours preferred) Start Date Is: ASAP Duration ... A Data Engineer will help build and maintain our data infrastructure. Building scalable data ...

Blockchain, Data Analytics (Remote)

New York, NY ยท On-site +1

$80K - $150K/yr

Develop statistical models and data science solutions while also implementing the infrastructure to ... Remote Work Environment * Maternity/Paternity Leave $80,000 - $150,000 a year The salary range for ...

next page

Showing results 1-20

Remote Data Infrastructure information

What are the key skills and qualifications needed to thrive as a Remote Data Infrastructure Engineer, and why are they important?

To excel as a Remote Data Infrastructure Engineer, you need a strong background in computer science, data architecture, and experience with cloud platforms such as AWS, Azure, or Google Cloud. Familiarity with tools like Terraform, Kubernetes, and data pipeline technologies, as well as relevant certifications (e.g., AWS Certified Solutions Architect), is typically required. Strong problem-solving abilities, clear communication, and self-motivation are essential soft skills for remote collaboration and troubleshooting. These competencies ensure reliable, scalable data systems and effective teamwork across distributed environments.

What are some common challenges faced by professionals working in remote data infrastructure roles?

Professionals in remote data infrastructure roles often encounter challenges such as ensuring seamless communication across distributed teams, maintaining high availability and performance of data systems, and managing security risks associated with remote access. Coordinating with colleagues across different time zones can require flexibility in scheduling and proactive communication. Additionally, remote data infrastructure engineers must stay up-to-date with evolving cloud technologies and best practices to effectively support scalable, reliable, and secure data architectures.

What is remote data infrastructure?

Remote data infrastructure refers to the systems, tools, and processes that enable organizations to collect, store, manage, and analyze data from remote locations, often via cloud-based platforms. This infrastructure allows teams to access and work with data securely from anywhere, supporting distributed work environments and scalable data solutions. It typically involves cloud storage, data pipelines, databases, and security protocols tailored for remote accessibility. Remote data infrastructure is essential for businesses that operate in multiple locations or have remote teams.
What are the most commonly searched types of Data Infrastructure jobs in New York? The most popular types of Data Infrastructure jobs in New York are:
What are popular job titles related to Remote Data Infrastructure jobs in New York? For Remote Data Infrastructure jobs in New York, the most frequently searched job titles are:
What job categories do people searching Remote Data Infrastructure jobs in New York look for? The top searched job categories for Remote Data Infrastructure jobs in New York are:
What cities in New York are hiring for Remote Data Infrastructure jobs? Cities in New York with the most Remote Data Infrastructure job openings:
Data/Infrastructure Advocate Engineer - US Remote

Data/Infrastructure Advocate Engineer - US Remote

Hugging Face

New York, NY โ€ข On-site, Remote

$125K - $150K/yr

Full-time

Medical, Dental, Vision, PTO

Posted 25 days ago


Job description

At Hugging Face, we're on a journey to democratize good AI. We are building the fastest growing platform for AI builders with over 11 million users who collectively shared over 4 million models, 1 million datasets & 1.5 million Gradio apps. Our open-source libraries have more than 700,000 stars on Github.
About the Role
As our first Data/Infrastructure Advocate Engineer, you'll bridge the gap between cutting-edge data infrastructure and the global community of data engineers, researchers, and developers. You'll champion Xet storage on the Hugging Face Hub, helping users efficiently store, version, and collaborate on large-scale datasets. This role is for someone who thrives at the intersection of technical depth (storage, Parquet, deduplication) and community advocacy, helping define the future of open data workflows.
You'll collaborate with teams like Datasets, Hub, and Infrastructure to shape how developers interact with data on our platform, and inspire a community to build better, faster, and more scalable data pipelines.
Your main missions
  • Grow and nurture the open-source data/infra community: launch initiatives, collaborate with data-focused groups, and organize events or challenges. Engage with communities like Apache Parquet, Open Table Formats, and data engineering forums to promote best practices and Hugging Face tools.
  • Promote the Hugging Face Hub as the go-to platform for data storage, versioning, and collaboration, curating and showcasing datasets, benchmarks, and tools like Xet.
  • Highlight use cases like efficient large-dataset updates, Parquet editing, and deduplication to demonstrate the Hub's value for data workflows.
  • Create demos, benchmarks, and tools (for example Colab notebooks) that illustrate best practices for data storage and versioning, and experiment with Xet, Parquet, and other formats.
  • Produce high-quality tutorials, blog posts, and videos that make complex topics accessible.
  • Share insights on storage optimization, dataset versioning, and deduplication to empower developers.
  • Actively participate in online communities (Discord, GitHub, forums) to highlight contributions, answer questions, and foster collaboration.
  • Make sure datasets and tools released on the Hub are well-documented, with clear examples, benchmarks, and use cases.
About You
You're already an active voice in the data and ML community. You build in public, you publish, and people follow your work on LinkedIn and X.
You're a hands-on builder who loves experimenting with data tools, storage optimization, and dataset versioning. You can take a complex topic like deduplication, compression, or Parquet editing and make it click for other developers through writing, demos, or talks. You're passionate about open source and knowledge sharing, and you thrive in fast-moving environments.
What you'll need
  • 3+ years in developer relations or developer advocacy, ideally for data engineering, infrastructure, or ML tools and platforms
  • An established public presence as a technical voice, with a track record of regularly publishing data/infra/ML content and a demonstrable, engaged audience on LinkedIn and X (Twitter)
  • A portfolio of developer-facing content you can point to: tutorials, blog posts, videos, demos, benchmarks, or conference talks
  • Hands-on experience building and engaging open-source or developer communities (Discord, GitHub, forums)
  • Strong Python skills
  • Hands-on experience with data libraries such as pandas, pyarrow, and huggingface/datasets
  • Practical experience with storage systems and formats: Parquet, Open Table Formats, and S3
  • Working knowledge of dataset versioning, deduplication, and compression
  • Ability to explain complex technical topics clearly through writing, demos, or talks
  • Fluent written and spoken English
Nice to have
  • Experience with the Hugging Face Hub and datasets ecosystem, or with Xet
  • Open-source maintainer or contributor experience
  • Familiarity with large-scale data pipelines and data engineering workflows
  • Experience producing notebooks (for example Colab) for tutorials and benchmarks
A note on fit
If you're interested in joining us but don't tick every box above, we still encourage you to apply. We're building a diverse team whose skills, experiences, and backgrounds complement one another, and we're happy to consider where you might make the biggest impact.
One more thing
At Hugging Face we believe great AI shouldn't require a massive cluster, we build for everyone, especially the GPU-poor. And because we read every application, here's a small sign that you read this one too: start your answer to the first application question with the words "GPU-poor and proud ". No trick, no catch, it just tells us a real person is on the other side.
More about Hugging Face
We are actively working to build a culture that values diversity, equity, and inclusivity. We are intentionally building a workplace where you feel respected and supported-regardless of who you are or where you come from. We believe this is foundational to building a great company and community, as well as the future of machine learning more broadly. Hugging Face is an equal opportunity employer, and we do not discriminate based on race, ethnicity, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or ability status.
We value development. You will work with some of the smartest people in our industry. We are an organization that has a bias for impact and is always challenging ourselves to grow continuously. We provide all employees with reimbursement for relevant conferences, training, and education.
We care about your well-being. We offer flexible working hours and remote options. We offer health, dental, and vision benefits for employees and their dependents. We also offer parental leave and flexible paid time off.
We support our employees wherever they are. While we have office spaces in NYC and Paris, we're very distributed, and all remote employees have the opportunity to visit our offices. If needed, we'll also outfit your workstation to ensure you succeed.
We want our teammates to be shareholders. All employees have company equity as part of their compensation package. If we succeed in becoming a category-defining platform in machine learning and artificial intelligence, everyone enjoys the upside.