1

Observability Internship Jobs in Texas (NOW HIRING)

Trinity Industry is looking for Data Analytics Interns for our office in Dallas, TX . This position ... observability, basic cost and drift checks). * Experience building or working with data/workflow ...

... and observability of our systems. * Work closely with internal partners to drive successful ... OR 3+ years of professional experience building full-stack software in lieu of a degree (internship ...

Sr. Software Engineer (Starlink)

Bastrop, TX

$121K - $160K/yr

... and observability of our systems. * Work closely with internal partners to drive successful ... OR 7+ years of professional experience building full-stack software in lieu of a degree (internship ...

Trinity Industry is looking for Data Analytics Interns for our office in Dallas, TX . This position ... observability, basic cost and drift checks). * Experience building or working with data/workflow ...

Lead Machine Learning Engineer

Plano, TX · On-site

$98K - $129K/yr

Build and integrate scalable evaluation (Evals) and observability frameworks into solutions to ... Internship experience does not apply) * At least 4 years of experience programming with Python ...

Lead Machine Learning Engineer

Plano, TX

$98K - $129K/yr

Build and integrate scalable evaluation (Evals) and observability frameworks into solutions to ... Internship experience does not apply) * At least 4 years of experience programming with Python ...

next page

Showing results 1-20

Observability Internship information

What are the key skills and qualifications needed to thrive as an Observability Intern, and why are they important?

To thrive as an Observability Intern, you generally need foundational knowledge in computer science, familiarity with monitoring concepts, and experience with programming or scripting languages. Exposure to observability tools like Prometheus, Grafana, ELK stack, or cloud monitoring platforms, along with coursework or certifications in DevOps or cloud technologies, is often beneficial. Strong analytical thinking, problem-solving abilities, and effective communication help interns collaborate with engineering teams and interpret complex data. These skills enable interns to contribute to system reliability by identifying, diagnosing, and resolving performance issues efficiently.

What is an Observability Internship?

An Observability Internship is a temporary, often entry-level position where students or recent graduates learn about and assist with monitoring, measuring, and analyzing the performance and reliability of software systems. Interns work with tools that provide visibility into applications, infrastructure, and services to help teams detect issues and improve system health. The role typically involves tasks such as setting up dashboards, analyzing logs and metrics, and helping to implement best practices for observability. This internship is valuable for those interested in DevOps, Site Reliability Engineering, or software development roles.

What is the difference between Observability Internship vs Monitoring Internship?

AspectObservability InternshipMonitoring Internship
FocusBroad system insights, including logs, metrics, tracesReal-time system health and alerting
SkillsData analysis, debugging, understanding of distributed systemsAlert configuration, basic system metrics
Work EnvironmentDevOps, SRE teams, cloud environmentsOperations, system administration teams
CertificationsKnowledge of monitoring tools (Prometheus, Grafana), scriptingBasic monitoring tools, scripting skills

While both internships involve system health, an Observability Internship covers a broader set of tools and concepts like logs, traces, and metrics for comprehensive system understanding. Monitoring internships focus more on real-time alerts and system uptime. Understanding these differences helps candidates choose the right role aligned with their skills and career goals.

What types of projects and tools can I expect to work with during an Observability Internship?

As an Observability Intern, you'll often contribute to projects focused on monitoring, logging, and tracing the performance of software systems. You may work with popular tools such as Prometheus, Grafana, ELK Stack (Elasticsearch, Logstash, Kibana), or OpenTelemetry to collect and visualize system metrics. Interns typically collaborate closely with site reliability engineers and software developers to identify bottlenecks, improve alerting, and ensure system reliability. This role provides hands-on experience with real-world infrastructure and fosters valuable problem-solving skills in a collaborative, technical environment.
What are the most commonly searched types of Observability jobs in Texas? The most popular types of Observability jobs in Texas are:
What cities in Texas are hiring for Observability Internship jobs? Cities in Texas with the most Observability Internship job openings:
Infographic showing various Observability Internship job openings in Texas as of July 2026, with employment types broken down into 2% Internship, 76% Full Time, 20% Part Time, 1% Temporary, and 1% Contract. Highlights an 87% Physical, 1% Hybrid, and 12% Remote job distribution.

Software Development Engineer, AWS OpenSearch Service

Amazon

Austin, TX • On-site

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

Re-posted 1 hour ago


Amazon rating

7.4

Company rating: 7.4 out of 10

Based on 7,025 frontline employees who took The Breakroom Quiz

6th of 39 rated national retailers


Job description

Imagine running search and analytics at any scale without thinking about clusters, capacity, or version upgrades. Where your queries return in milliseconds whether you're indexing a thousand documents or a hundred billion. Where the platform scales, heals, and secures itself so your team can focus on the questions, not the infrastructure.
That's what we build.
Amazon OpenSearch Service is the fully managed and Serverless AWS service that lets customers deploy, operate, and scale OpenSearch for log analytics, full-text search, application monitoring, observability, and AI-powered retrieval. We serve hundreds of thousands of customers running mission-critical workloads across every AWS region, and we are part of the AWS Database and Analytics organization.
Come join the OpenSearch Control Plane team. We are responsible for a service that:
- Reliably manages a large fleet of cloud-native OpenSearch clusters and collections, freeing customers from sizing, scaling, and patching decisions
- Guarantees high availability and durability for mission-critical search and observability workloads at the scale of the most demanding internet businesses
- Orchestrates and automates the complete lifecycle of an OpenSearch cluster or collection - from creation through scale-up, scale-out, upgrade, replication, and fail-over
The OpenSearch control plane is not just any distributed system. It orchestrates a fleet of clusters across every AWS region, detects and recovers from node failures in seconds, and serves workloads that span petabytes of customer data. Most recently, we launched **OpenSearch Serverless NextGen** - delivering ultra-fast collection provisioning and true scale-to-zero capacity, so customers pay only for what they use and get a working environment in seconds rather than minutes. We are not done; this is one of many investments raising the bar for what customers can expect from a managed search service.
OpenSearch itself is a community-driven, Apache 2.0-licensed open-source search and analytics suite built on Apache Lucene. Since launching in July 2021, the project has delivered multiple major and minor releases advancing core search, vector search, and analytics. Our next phase focuses on transforming OpenSearch into an intelligent retrieval engine - one that meets the evolving needs of customers running search, analytics, and AI-powered workloads across structured, unstructured, and multi-modal data.
We work in Java, Rust, Python, and Go. We contribute upstream to OpenSearch and Apache Lucene. We are customer-obsessed, operate with the entrepreneurial pace of a startup inside one of the world's largest cloud providers, and value strong intuition backed by metrics.
**Operations is core engineering on this team, not overhead.** Engineers operate what they build, and we treat customer-impacting incidents with the same engineering rigor we apply to feature work. Reliability and Static Stability is a forcing function for design, not a phase after launch.
Key job responsibilities
- Design, develop, and operate components of a large-scale distributed control plane - reason about failure modes, multi-tenant isolation, and blast-radius before writing code
- Look for opportunities to simplify - consolidate duplicate abstractions, remove dead code, reduce configuration sprawl, and challenge complexity that doesn't earn its keep
- Treat operations as core engineering - own the systems you build end-to-end, participate fully in incident response and RCAs, and drive operational improvements that reduce cost and maintenance burden
- Solve real problems in cluster orchestration, capacity provisioning, lifecycle workflows, multi-tenant infrastructure, query routing, and operational tooling
- Use AI coding assistants and GenAI tools effectively across the software development lifecycle - from design and prototyping through coding, testing, code review, debugging, and operations - applying engineering judgment to validate AI-generated output against Amazon's correctness, security, and operational bar
- Write robust, efficient, and maintainable code in Java, Rust, Python, or Go
- Partner with senior engineers to translate ambiguous requirements into concrete deliverables; prototype, test, and validate
- Continually challenge what exists and explore what should change to better serve customers
BASIC QUALIFICATIONS
- 3+ years of non-internship professional software development experience
- 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience
- 1+ years of software development engineer or related occupational experience
- 1+ years of designing and developing large-scale, multi-tiered, multi-threaded, embedded or distributed software applications, tools, systems, and services using: C#, C++, Java, or Perl experience
- 1+ years of Object Oriented Design experience
- Bachelor's degree or foreign equivalent in Computer Science, Engineering, Mathematics, or a related field
- Experience programming with at least one software programming language
PREFERRED QUALIFICATIONS
- 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience
- Bachelor's degree in computer science or equivalent
Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.
Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner.
The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.
USA, TX, Austin - 143,700.00 - 194,400.00 USD annually

What Amazon employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Amazon logo

About Amazon

Sourced by ZipRecruiter

Amazon.com, Inc., commonly known as Amazon, is an American multinational technology company. It was founded by Jeff Bezos in 1994 and initially started as an online marketplace for books. Since then, Amazon has expanded its operations and become one of the largest e-commerce companies in the world. Amazon's primary business is its online retail platform, where customers can purchase a vast array of products, including electronics, clothing, books, home goods, and much more. The company offers a convenient and user-friendly shopping experience, with features such as fast shipping, customer reviews, and personalized recommendations. In addition to its e-commerce platform, Amazon has diversified its business into various other areas. One of its notable ventures is Amazon Web Services (AWS), a comprehensive cloud computing platform that provides services such as storage, compute power, and database management to individuals and businesses. AWS has become a leader in the cloud computing industry, powering many websites and applications worldwide. Amazon has also developed its own consumer electronics, including the popular Amazon Kindle e-reader, Fire tablets, Fire TV streaming devices, and the Alexa-powered Echo smart speakers. The Alexa voice assistant, integrated into these devices, allows users to interact with their devices using voice commands, perform tasks, and access information. Furthermore, Amazon has expanded into media and entertainment. It operates Prime Video, a streaming service that offers a wide range of movies, TV shows, and original content. Amazon Music provides a platform for streaming and purchasing digital music, while Audible offers audiobooks and other audio content. The company's commitment to customer satisfaction and convenience is demonstrated by its membership program, Amazon Prime. Prime members receive various benefits, including free two-day shipping, access to streaming services, exclusive deals, and more.

Industry

It services, book publishers, retail, real estate and computer and electronic product manufacturing

Company size

10,000+ Employees

Headquarters location

Seattle, WA, US