Skip to main content

Multimodal AI Systems Architect (AI Engineering)

Expired
This role has expired and is no longer accepting applications. Browse similar roles →
Hyphen Connect Limited
Australia
remote
Full Time / Permanent

Apply for this job

Posted 1 month ago
This role is expired

These roles are hiring now

View all similar roles →

AI Research Engineer (Kernel & Inference Optimization)

Jobgether
Australia
remote
  • AI model-serving architecture, inference optimization & GPU kernel development
  • PhD in NLP/ML highly relevant, strong AI research track record required
  • Metal Shading Language, GPU kernels, quantization, distributed inference
Posted 5d ago

Principal Adviser Data & AI Architecture

Rio Tinto
Brisbane, QLD | Perth, WA
hybrid
  • Define enterprise architecture for Data & AI platforms globally
  • Deep expertise in Databricks, Lakehouse architecture, Unity Catalog
  • Multi-cloud architecture, TOGAF, governance, responsible AI frameworks
Posted 4d ago

Senior Forward Deployed Engineer, GenAI, Google Cloud

Google
Sydney, NSW
  • Build production-grade GenAI solutions, multi-agent systems, agentic workflows
  • 5+ years software development (Python), AI systems experience required
  • RAG architectures, LLMs, vector databases, GCP, Vertex AI
Posted 4d ago

Applied AI Architect

OpenAI
Sydney, NSW
hybrid
  • Senior technical owner for customer AI strategy and production deployment
  • Significant experience in solutions architecture or technical account leadership
  • Enterprise AI systems, LLMs, cloud architecture, Python/JavaScript, APIs
Posted 4d ago

We are seeking a talented Multimodal AI Systems Architect to develop and optimize AI systems that seamlessly integrate vision and audio models. This role focuses on enhancing our voice-to-voice interactions and multimodal retrieval capabilities, ensuring our systems are efficient and innovative.

Responsibilities:

  • Integrate vision encoders and audio-native models into core agent reasoning loops.
  • Optimize streaming latency for voice-to-voice AI interactions.
  • Architect multimodal RAG systems capable of retrieving insights from videos and PDFs.

Qualifications:

  • Experience with Whisper, CLIP, and multimodal LLM integration.
  • Knowledge of streaming architectures and WebRTC.
  • Expertise in cross-modal alignment.