IT Site Reliability Engineer — API Management Platforms

Socket.dev

Dallas (TX)

On-site

USD 140,000 - 190,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Texas Instruments is seeking an IT Site Reliability Engineer to own TI's Apigee Edge private cloud platform and provide SRE support for additional automation and DevOps tooling. You will collaborate with security, network, and infrastructure teams to ensure platforms are stable, secure, and scalable across dev/test and production environments.

This role sits at the intersection of infrastructure administration, DevOps engineering, and platform reliability, requiring deep API management

Responsibilities

  • Apigee Platform Ownership: own installation, configuration, operations, upgrades and roadmap evolution of Apigee Edge private cloud.
  • Platform Administration: install, configure, upgrade, patch, and maintain Apigee Edge components across dev/test and production.
  • Monitoring & Intervention: own monitoring to detect issues early and implement self-healing where possible.
  • Incident Management: act as primary escalation for Apigee incidents; lead RCA and PIRs.
  • API Proxy Lifecycle Support: assist with API proxy pipelines, policy troubleshooting, and runtime performance.
  • Security & Compliance: enforce security hardening, manage TLS/SSL, OAuth and RBAC per TI policies.
  • Capacity Planning & Performance Tuning: monitor resources and tune performance across components.
  • Backup & Disaster Recovery: maintain backup/restore and DR procedures for Apigee data stores.
  • Vendor Engagement: manage relationship with Google/Apigee support for issues and roadmaps.
  • Platform Roadmap Input: advise on Apigee strategy including potential migration paths.
  • Multi-Platform SRE Support: provide SRE for CI/CD tools, artifact management, source control, and tooling.
  • Operational Consistency: define SLOs/SLIs, dashboards, runbooks, and reliability improvements.
  • Cross-Platform Incident Support: triage incidents across platforms with SMEs.
  • Automation & Tooling Development: build IaC automation scripts (Python, Bash, Ansible, Terraform).
  • Documentation & Knowledge Management: maintain architecture docs and runbooks for all platforms.
  • Cross-Team Collaboration: partner with networking, security and infra teams for integrations and improvements.

Job description

About the Role

The Enterprise Platforms team within our Data & Agentic Platform Solutions organization is responsible for deploying, operating, and continuously improving the platforms that power TI's digital integration, automation and DevOps capabilities. As an IT Site Reliability Engineer within the Enterprise Platforms team, you will serve as the primary technical platform owner for TI's Apigee Edge private cloud environment — the backbone of TI's API management and integration strategy — while also providing platform SRE support across a broader portfolio of automation and DevOps tooling used throughout the enterprise.


This role sits at the intersection of infrastructure administration, DevOps engineering, and platform reliability. You will be expected to bring deep technical expertise in API management operations while also developing familiarity with adjacent platforms such as CI/CD pipelines, artifact repositories, source control systems, automation orchestration tools, and integration middleware. You will work closely with application development teams, security, network, and infrastructure teams to ensure all supported platforms are stable, performant, secure, and scalable.


Key Responsibilities

Apigee Platform Ownership (Primary Focus)


  • Primary Technical Owner: Serve as the go-to subject matter expert (SME) for TI's Apigee Edge private cloud platform, owning the full platform lifecycle from installation and configuration through ongoing operations, upgrades, and eventual roadmap evolution.

  • Platform Administration: Install, configure, upgrade, patch, and maintain Apigee Edge private cloud components (Management Server, Router, Message Processor, Cassandra, ZooKeeper, Qpid, Postgres) across dev/test and production environments.

  • Monitoring & Intervention: Own and improve system monitoring solution to ensure early detection of issues and implementation of self-healing solutions where appropriate.

  • Incident Management: Serve as the primary escalation point for all Apigee-related incidents; lead root-cause analysis (RCA) and drive post-incident reviews (PIRs) to prevent recurrence.

  • API Proxy Lifecycle Support: Partner with development teams on API proxy deployment pipelines, policy troubleshooting, and runtime performance optimization within the Apigee environment.

  • Security & Compliance: Apply and enforce security hardening standards on the Apigee platform; manage TLS/SSL certificates, OAuth configurations, and role-based access controls (RBAC); ensure alignment with TI IT security policies and audit requirements.

  • Capacity Planning & Performance Tuning: Monitor platform resource utilization (CPU, memory, disk, JVM heap); conduct capacity planning and performance tuning across all Apigee Edge components and underlying infrastructure.

  • Backup & Disaster Recovery: Own and maintain backup, restore, and disaster recovery procedures for all Apigee Edge components and associated data stores (Cassandra, PostgreSQL, ZooKeeper).

  • Vendor Engagement: Manage the technical relationship with Google/Apigee support for escalated issues, product defects, version roadmaps, and advisory services.

  • Platform Roadmap Input: Provide technical guidance and recommendations on Apigee platform strategy, including future-state considerations such as migration paths to Apigee Hybrid hosted solution.


Broader Automation & DevOps Platform SRE Support (Secondary Focus)


  • Multi-Platform SRE: Provide site reliability engineering support for other automation and DevOps platforms within the ITS DevOps Platforms portfolio — which may include CI/CD tools, artifact management, source control, integration middleware, and automation orchestration platforms.

  • Operational Consistency: Apply consistent SRE principles across all supported platforms — defining SLOs/SLIs, building observability dashboards, maintaining runbooks, and driving reliability improvements.

  • Cross-Platform Incident Support: Respond to and triage incidents across the broader platform portfolio; coordinate with platform-specific SMEs and infrastructure teams to restore services rapidly.

  • Automation & Tooling Development: Develop and maintain automation scripts and Infrastructure-as-Code (IaC) tooling (Python, Bash, Ansible, Terraform, or similar) to streamline operations, deployments, and configuration management across all supported platforms.

  • Documentation & Knowledge Management: Create and maintain thorough technical documentation including architecture diagrams, operational runbooks, change records, and knowledge base articles for all platforms under support.

  • Cross-Team Collaboration: Partner with networking, server infrastructure, application development, and security teams to support integrations, troubleshoot complex issues, and deliver platform enhancements across the automation portfolio.


Why TI?


  • Engineer your future. We empower our employees to truly own their career and development. Come collaborate with some of the smartest people in the world to shape the future of electronics.

  • We're different by design. Diverse backgrounds and perspectives are what push innovation forward and what make TI stronger. We value each and every voice, and look forward to hearing yours.

  • Benefits that benefit you. We offer competitive pay and benefits designed to help you and your family live your best life. Your well-being is important to us. Please find our country-specific benefits here


About Texas Instruments

Texas Instruments Incorporated (Nasdaq: TXN) is a global semiconductor company that designs, manufactures and sells analog and embedded processing chips for markets such as industrial, automotive, data center, personal electronics and communications equipment. At our core, we have a passion to create a better world by making electronics more affordable through semiconductors. This passion is alive today as each generation of innovation builds upon the last to make our technology more reliable, more affordable and lower power, making it possible for semiconductors to go into electronics everywhere. Learn moreatTI.com.


Texas Instruments is an equal opportunity employer and supports a diverse, inclusive work environment. All qualified applicants will receive consideration for employment without regard to race, color, religion, creed, disability, genetic information, national origin, gender, gender identity and expression, age, sexual orientation, marital status, veteran status, or any other characteristic protected by federal, state, or local laws.


TI does not make recruiting or hiring decisions based on citizenship, immigration status or national origin. However, if TI determines that information access or export control restrictions based upon applicable laws and regulations would prohibit you from working in this position without first obtaining an export license, TI expressly reserves the right not to seek such a license for you and either offer you a different position that does not require an export license or decline to move forward with your employment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

IT Site Reliability Engineer — API Management Platforms
IT Site Reliability Engineer — API Management Platforms

Texas Instruments Incorporated • Dallas (TX), Northern (KY)

Hybrid
USD 140,000 - 190,000
IT Site Reliability Engineer — API Management Platforms
IT Site Reliability Engineer — API Management Platforms

Texas Instruments • Dallas (TX), Northern (KY)

Hybrid
USD 120,000 - 180,000
IT Site Reliability Engineer — API Management Platforms
IT Site Reliability Engineer — API Management Platforms

118-WW TMG MFG OPS • Dallas (TX)

On-site
USD 130,000 - 180,000
Product Engineering Intern
Product Engineering Intern

Texas Instruments • Knoxville (TN)

On-site
USD 28,000 - 41,000
Data Scientist
Data Scientist

Socket.dev • Dallas (TX)

On-site
USD 140,000 - 210,000
System Engineer | INT PD
System Engineer | INT PD

Texas Instruments • Dallas (TX)

On-site
USD 100,000 - 150,000
Applications Engineering Intern
Applications Engineering Intern

Texas Instruments • Knoxville (TN)

On-site
USD 42,000 - 52,000
Product Operations Analyst - Encore Program Internship
Product Operations Analyst - Encore Program Internship

Texas Instruments • Dallas (TX)

On-site
USD 24,796 - 30,307
API Platform SRE: Apigee Edge & DevOps Owner
API Platform SRE: Apigee Edge & DevOps Owner

Socket.dev • Dallas (TX)

On-site
USD 140,000 - 190,000
Network Engineer Encore Program Internship
Network Engineer Encore Program Internship

Socket.dev • Dallas (TX)

On-site