Skip to main content

Multimodal AI Systems Architect (AI Engineering)

Expired
This role has expired and is no longer accepting applications. Browse similar roles →
Hyphen Connect Limited
Australia
remote
Full Time / Permanent

Apply for this job

Posted 26d ago
This role is expired

These roles are hiring now

View all similar roles →

Forward Deployed AI Engineer

Cloudera
Sydney, NSW
remote
  • Build and deploy AI apps, embed with customers, shape enterprise AI strategy
  • 5+ years engineering experience, 2-3 years in Generative AI/agentic apps
  • LLMOps, RAG, foundation models, full-stack AI, enterprise architecture
Posted 5d ago

Senior AI/ML Architect - LLM & Agentic AI

Minfy
Australia
remote
  • Solution architecture & pre-sales for LLM, Agentic AI & AWS Bedrock
  • 12-18 years experience in LLM & Agentic AI
  • AWS Bedrock, Claude LLMs, RAG pipelines, Python, prompt engineering
Posted 9d ago

Senior Advisory AI Foundry Architect

ServiceNow
Melbourne, VIC
hybrid
  • Design and implement AI solutions on ServiceNow AI platform
  • 8+ years enterprise architecture experience
  • AI/ML, ServiceNow platform, enterprise architecture, solution design
Posted 7min ago

Applied AI Architect, Partnerships

Anthropic
Sydney, NSW
hybrid
  • Pre-sales architect for GSI, RSI & cloud partnerships
  • 10+ years in Solutions Architect, Sales Engineer or Partner Sales roles
  • Cloud architectures, LLM frameworks, GSI/cloud provider partnerships
Posted 22h ago

We are seeking a talented Multimodal AI Systems Architect to develop and optimize AI systems that seamlessly integrate vision and audio models. This role focuses on enhancing our voice-to-voice interactions and multimodal retrieval capabilities, ensuring our systems are efficient and innovative.

Responsibilities:

  • Integrate vision encoders and audio-native models into core agent reasoning loops.
  • Optimize streaming latency for voice-to-voice AI interactions.
  • Architect multimodal RAG systems capable of retrieving insights from videos and PDFs.

Qualifications:

  • Experience with Whisper, CLIP, and multimodal LLM integration.
  • Knowledge of streaming architectures and WebRTC.
  • Expertise in cross-modal alignment.