Technical Operations Engineer
RemoteWarsaw Metropolitan AreaseniorFull-time
- Posted
- today
- Source
- LinkedIn (remote, Europe)
- Field
- Engineering, Operations
Skills
PostgreSQLTerraformPythonDockerCI/CDLinuxNext.jsSQLAWSGitGoAI
Description
About InspektAI
InspektAI builds AI and computer-vision software for drone-based building inspections. Our SaaS platform runs almost entirely on AWS and serves customers across multiple global markets, with teams in Sweden, Thailand, Spain and the Philippines.
The role
- You will be the deep technical escalation point for our production platform: our own applications running on Linux and AWS. When something breaks and the runbook doesn't cover it, you dig through application logs, the OS and the database to find the root cause, fix it, and make sure it doesn't happen again.
- You will work alongside our senior AWS engineers: we need someone who understands the platform deeply, not someone who follows a checklist.
What you will do
- Own escalated incidents (3rd line). Take over issues our internal support team cannot resolve, drive them to root cause, and fix them properly rather than patching symptoms.
- Troubleshoot our own applications. Find root causes in our self-developed services running on AWS by working through logs, metrics, configuration and code, and work with developers to get fixes shipped.
- Run and tune Linux. Keep our Linux hosts and containers healthy, secure and performant, and debug at the OS level when the application layer doesn't explain the problem.
- Support the database. Investigate slow queries, locking, connection issues and capacity on RDS PostgreSQL.
- Share on-call for production with our senior AWS engineer and act as their backup.
- Operate our AWS environment (Elastic Beanstalk, AWS Batch, Lambda, EC2, Fargate, S3, EFS, CloudFront, SQS, CloudWatch and more), with changes made in Terraform and reviewed in GitHub.
- Improve observability. Tune monitoring, alerting and logging so problems are caught early and alerts mean something.
- Close the loop. Write postmortems and keep SOPs and runbooks accurate, so the internal support team can handle recurring issues without escalating. Help level up their technical depth.
- Keep things secure and efficient: IAM least privilege, patching, and an eye on AWS cost.
What we are looking for
- 2-3+ years operating production workloads on AWS.
- Deep Linux knowledge: processes, memory, filesystems, systemd, networking and performance debugging from the command line (strace, lsof, tcpdump and similar).
- A track record of troubleshooting applications in production: reading logs, stack traces and metrics to find root causes in software you didn't write.
- Solid SQL and PostgreSQL skills, ideally on RDS: query analysis with EXPLAIN, indexing, locking and connection management.
- Hands-on experience with the AWS services listed above, and an understanding of how they behave and fail.
- Working experience with Terraform.
- Scripting in Bash and Python.
- Solid networking knowledge: VPC design, DNS, TLS, load balancing.
- Containers (Docker, ECS/Fargate) and Git-based CI/CD workflows.
- Clear written and spoken English. You document as you go.
Nice to have: running compute-heavy batch or data pipelines at scale (e.g. imagery or 3D processing).
Who you are
- An owner. You take a problem from "something is wrong" to "fixed, documented and prevented" without being chased.
- Self-driven. You don't wait for training or certificates. You read the docs, the source and the error logs, and you experiment until you understand.
- Genuinely into this tech. You enjoy digging into why a system behaves the way it does.
- A team player, not a cowboy. Production changes go through review, you communicate what you are doing, and you leave things clearer for the next person.
What we offer
- Competitive salary and benefits. [Range: TBD]
- Real ownership of a production cloud platform.
- An international team and flexible, remote-friendly working.
- Room to grow as the company scales.
Interested? Apply with your CV and a short note about a production problem you tracked down and what you learned from it.
JobMatch aggregates public listings. Always apply through the original posting.