寻找你的下一个职业机会

按职位、技能和地点搜索招聘信息。准备申请前,先仔细了解职位要求。

找到一个吸引人的职位名称只是求职的起点。请将工作职责、招聘要求和工作条件与你的实际经历进行比较。本指南帮助你筛选机会、准备有针对性的申请材料,并确认每份申请应该在哪里提交。

此界面为简体中文。雇主发布的职位名称和描述保留原文,可能为英文。

清除筛选

搜索结果: 7,991

← 返回搜索结果

Site Reliability Engineer

Intermedia Intelligent Communications

远程工作

地点
Remote (Portugal)
工作安排
Full Time
发布日期
2026年10月8日

申请前,请在雇主网站确认职位仍在招聘,并检查完整要求和条件。

职位描述

此界面为简体中文。雇主发布的职位名称和描述保留原文,可能为英文。

*ALL CANDIDATES MUST BE LOCATED IN PORTUGAL* About IntermediaAre you looking for a company where YOUR VOICE is heard? Where can you MAKE A DIFFERENCE? Do you THRIVE in a FAST-PACED work environment? Do you wake every morning EXCITED to work with GREAT PEOPLE and create SUCCESS TOGETHER? Then Intermedia is the place for you. Intermedia has established itself as a leading provider of cloud communications and collaboration tech that allows companies to connect better. We have a strong track record of growth, profitability, and creating an environment where everyone matters. Everyone. While we are fast-paced and admittedly a bit intense, we promise that you won’t be bored. You will find Intermedia is a place where you can indulge your passion for creating and supporting great cloud technology. What’s more, we always look to promote from within and have many employees who have been with us 10, 15, and 20+ years! Culture at Intermedia is built on teamwork and transparency. We hold each other accountable and always have each other’s back!While primarily remote, this role requires occasional visits to the office in Coimbra or in Aveiro. We plan to open an office in Porto in the future. This approach gives team members the flexibility to work remotely while also coming together in the office for collaboration and teamwork. Are you ready to make your mark? About the RoleWe are looking for a Site Reliability Engineer (SRE) to improve the reliability of our AI and analytics platforms and the services that depend on them. As Intermedia expands its global product deployments, this role will strengthen SRE practices across production infrastructure, data pipelines, machine-learning services, customer-facing analytics, and AI-powered Voice and Unified Communications capabilities. You will partner with AI, data, product, and platform teams to build scalable, observable systems that deliver dependable insights and resilient customer experiences. What you will be doing: Run and improve production environments that support AI and AI workloads, data pipelines, analytics applications, and customer-facing services. Build software and automation to manage cloud infrastructure, data platforms, model-serving infrastructure, and application services. Define and measure service level indicators, service level objectives, and error budgets for AI and analytics services, including availability, data freshness, pipeline completion, and inference latency. Build end-to-end observability that correlates metrics, logs, and traces with data-quality signals, AI-service performance, and customer impact. Monitor and optimize the reliability, performance, capacity, and cost of batch and streaming workloads, analytics queries, and inference services. Partner with data engineering and machine-learning teams to make ingestion, transformation, feature, training, deployment, and reporting workflows production-ready. Automate CI/CD and production-readiness checks for data pipelines, model and prompt releases, schema changes, analytics applications, and dashboards. Detect and resolve data-quality incidents involving missing, stale, delayed, or anomalous data, schema drift, and broken lineage or dependencies. Design and test graceful degradation, dependency isolation, retry and fallback patterns, and recovery procedures for impaired AI, data, or downstream services. Improve the reliability of AI-powered Voice and Unified Communications capabilities such as speech recognition, transcription, summarization, intelligent routing, conversational assistance, and text-to-speech. Establish operational monitoring for model and AI-service behavior, including latency, throughput, error rates, drift indicators, and changes in output quality. Plan capacity and run performance, load, and resilience tests across compute-intensive AI workloads, distributed data processing, and analytics services. Lead incident response and post-incident improvement for AI and analytics services, using measurable actions to reduce recurrence and recovery time. Support secure and reliable access to cloud storage, processing, and query services used by analytics products and internal decision-making. Reduce operational toil and improve engineering productivity through platform tooling, runbooks, self-service automation, and clear operational standards. What you will bring to the role: Bachelor's degree in computer science, data engineering, software engineering, or another technical or scientific discipline, or equivalent practical experience. 4-7 years of experience in production operations, systems engineering, SRE or DevOps, software deployment, and maintenance of distributed production systems. Experience with cloud infrastructure, containers, Kubernetes, distributed systems, and scalable compute and storage services. Experience operating data processing, orchestration, storage, or analytics technologies, such as Kafka, Spark, Airflow, dbt, data warehouses, or comparable cloud services. Ability to use metrics, logs, traces, data-quality checks, freshness indicators, lineage, and service-level indicators to diagnose complex production issues. Experience with CI/CD, DataOps or MLOps practices, automated testing, controlled rollout, and rollback of data and AI service changes. Strong troubleshooting skills across Linux, applications, networks, APIs, data pipelines, and distributed service dependencies, with attention to security and access controls. Strong analytical problem-solving and cross-functional communication skills, with a proactive approach to reliability, performance, and continuous improvement. Diversity Inclusion and Equal Opportunity We hire, promote, and compensate employees based on their ability to perform their job responsibilities, without regard to race, color, creed, religion, sex, gender, marital status, national origin, ancestry, age, citizenship, physical or mental disability, sexual orientation, or any other basis protected by applicable law (collectively referred to in our Code of Conduct as “Protected Classes”). We do not tolerate employment discrimination in the workplace, and we are committed to making reasonable accommodations for identified disabilities or other limitations as required by all applicable laws. We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status. Originally posted on Himalayas

查找职位、比较要求,再准备申请

从你希望从事的职位或运用的技能开始搜索。调整地点和职业筛选,打开职位比较工作职责。如果没有结果,可以使用更短的关键词,或逐一移除筛选条件。

区分必备要求和优先条件,检查已列出的工作安排、薪资和地点。远程职位也可能要求特定居住国家、工作许可或工作时间重合。请向雇主确认职位信息和完整条件。

选择能够回应职位要求的真实经历,说明你的贡献,只使用能够证实的数字。遵循雇主的申请说明,并在提交前检查联系方式、文档内容和 PDF。

提交申请前的检查清单

求职常见问题

为什么有些职位使用英文?

职位名称和描述由招聘企业撰写。为避免改变招聘要求或工作条件,我们保留原文。操作界面和本指南使用简体中文。如果中文搜索没有结果,可以尝试使用职位发布语言中的名称或技能,例如“software engineer”。搜索词不会自动翻译,界面语言也不代表企业要求的申请语言。

远程职位是否允许从任何国家工作?

不一定。企业可能对居住国家、工作许可或工作时段有要求。请查看原始招聘页面中的具体条件。如果未说明,应先向企业确认,再判断能否从你所在的地区工作。“远程”标签本身并不代表没有地点限制。

搜索没有结果时应该怎么办?

尝试更通用的职位名称或单个技能,并逐一移除筛选条件。不同企业可能用不同名称描述相似工作。如果某个职位已经消失,请在企业招聘页面搜索其职位编号。扩大搜索范围不会让已经关闭的职位重新开放。

申请会通过 ResumizeAI 直接提交吗?

申请按钮会打开外部网站。请按照企业或招聘服务的说明,在该网站完成并确认提交。在 ResumizeAI 中准备简历并不等于已经申请职位。如果链接只打开企业网站,请先找到对应职位,再继续申请流程。

如何针对职位调整简历和求职信?

将招聘要求与能够解释清楚的项目、任务和成果联系起来。突出相关经历,不要添加未经实际掌握的技能或虚构成绩。在求职信中用具体例子说明申请动机,并遵守企业要求的语言和文件格式。提交前检查两份文件,确保内容准确、联系方式正确、链接可用。