Software Engineer – AI & Platform (CIT8112026)

 Posted an hour ago
     
2-5 years experience
Apply Now

Please mention DailyRemote when applying

AI Summary

You will design, develop, and maintain AI-powered applications and autonomous agents while managing cloud infrastructure and production systems. Additionally, you will participate in incident response, ensure system observability, and collaborate with internal teams to optimize engineering workflows.

About Us 

Buddle was founded to address the growing demand for reliable and efficient staffing solutions across diverse industries globally. Our mission is to seamlessly connect businesses with skilled professionals, fostering long-term partnerships that benefit both parties. At Buddle, we prioritise trust, value, and sustainability in every staffing solution we provide.


About the Role
We are looking for a highly capable, self-driven, and technically versatile Software Engineer to join our client's growing technology team.


This role combines software engineering, AI and agent development, cloud infrastructure, production support, and system observability. You will work across the technology stack to build and maintain AI-powered applications, internal platforms, and automated workflows while supporting the reliability and performance of production systems.

The ideal candidate is comfortable working across different technical areas, can troubleshoot and respond to production issues, and is able to work independently in a fully remote environment with minimal supervision.


If you enjoy solving complex technical problems, learning new technologies, and building reliable and scalable systems, we'd love to hear from you!


Key Responsibilities
AI & Agent Development

  • Design, develop, maintain, and improve AI-powered applications, autonomous agents, and workflow automation.
  • Build and maintain systems that generate content strategies and dispatch tasks into AI/agent workflows.
  • Work with LLM orchestration, prompt engineering, agent frameworks, and human-in-the-loop patterns.
  • Develop reliable workflows that integrate AI capabilities with internal systems and third-party platforms.
  • Evaluate and improve AI-generated outputs for accuracy, quality, reliability, and consistency.
  • Explore and implement improvements to AI-powered workflows as the team's technology and business requirements evolve.

Software Engineering & Application Development

  • Design, develop, test, deploy, and maintain software applications and internal engineering solutions.
  • Develop integrations with internal and third-party platforms using APIs and other integration technologies.
  • Work across Python and TypeScript codebases and contribute to system architecture and technical design.
  • Write clean, maintainable, and reliable code following established engineering standards and practices.
  • Identify opportunities to automate repetitive processes and improve engineering and operational efficiency.
  • Collaborate with internal teams to understand requirements and translate business needs into effective technical solutions.

Cloud Infrastructure & DevOps

  • Develop and maintain cloud infrastructure and application environments using AWS and Terraform.
  • Work with AWS services including ECS, SQS, Aurora, S3, and CloudFront.
  • Support infrastructure configuration, deployment, scaling, and operational maintenance.
  • Develop and maintain infrastructure-as-code using Terraform.
  • Assist with system configuration, security, credentials, and access management as required.
  • Contribute to infrastructure improvements that enhance reliability, scalability, and operational efficiency.

Production Support & Incident Management

  • Participate in the engineering on-call rotation and provide timely support for production incidents.
  • Monitor and triage application and infrastructure alerts using tools such as Datadog and Sentry.
  • Investigate and troubleshoot production issues, including database, payment, integration, synchronisation, and infrastructure-related incidents.
  • Stabilise production systems during incidents, investigate root causes, and implement or recommend appropriate resolutions.
  • Escalate complex issues when required while providing clear technical findings and relevant context.
  • Contribute to incident response procedures, playbooks, and post-incident documentation.
  • Identify recurring issues and implement improvements to reduce the likelihood of future incidents.

Observability & System Reliability

  • Develop and maintain effective monitoring, alerting, and observability across applications and infrastructure.
  • Create and improve Datadog monitors, dashboards, alerts, and operational visibility.
  • Identify gaps in monitoring coverage and implement appropriate alerts for critical system events.
  • Support observability across areas such as application failures, message queues, infrastructure events, database issues, and system performance.
  • Ensure new features and systems are designed with monitorability and operational reliability in mind.
  • Work proactively to identify potential system issues before they significantly impact production.

Data & Engineering Operations

  • Support Databricks pipelines, dashboards, and data-related workflows.
  • Investigate and resolve issues affecting data pipelines and system integrations.
  • Work with engineering and operational data to support troubleshooting, monitoring, and business processes.
  • Maintain accurate technical documentation and contribute to knowledge-sharing across the engineering team.


Other Responsibilities

  • Make recommendations for technical, operational, and process improvements that improve system reliability and team efficiency.
  • Maintain clear documentation of technical decisions, processes, configurations, and incident resolutions.
  • Collaborate with other engineers and internal stakeholders to deliver technical solutions effectively.
  • Stay current with emerging technologies, particularly in AI, agent frameworks, cloud infrastructure, and software engineering.
  • Contribute to a positive engineering culture through knowledge sharing, collaboration, and continuous improvement.
  • Perform other related tasks and duties that may be reasonably assigned by the client.

About You

  • 3+ years of experience in software engineering, application development, platform engineering, or a similar technical role.
  • Strong experience with Python and/or TypeScript and modern software development practices.
  • Hands-on experience developing AI-powered applications, LLM workflows, autonomous agents, or AI-driven automation.
  • Practical understanding of LLM orchestration, prompt engineering, and agent-based workflows.
  • Experience working with AWS or similar cloud platforms and Terraform or other Infrastructure as Code (IaC) tools.
  • Experience with production support, incident response, monitoring, and observability.
  • Experience working with APIs, integrations, and cloud-based applications.
  • Strong analytical and troubleshooting skills, with the ability to investigate and resolve complex technical issues.
  • Comfortable working across different areas of the technology stack and learning new technologies as required.
  • Highly self-driven, organised, and comfortable working independently in a fully remote environment with minimal supervision.
  • Strong written and verbal communication skills, with a proactive approach to documentation and continuous improvement.


Buddle Benefits Included

Health insurance

Internet allowance

KPI incentive program


Fortnightly virtual happy hour 

Annual group offsites


Finer Details

Schedule: TBD

Start Date:
TBD

Similar Jobs

See all Remote Software Development jobs →

Personalize your Remote Job Search in 3 Easy Steps!

Discover remote opportunities in Software Engineer

Answer easy questions

Answer easy questions

200,000+ jobs across 15+ categories

Get your best job matches

Get your best job matches

Only hand-screened, legit jobs

Find a remote job faster

Find a remote job faster

No ads, scams, or junk

I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!

Sarah J. — Sarah J. · Marketing Manager ★★★★★ Verified