AI Engineer

Cubeยท Data, Product & Technology
Apply Now โ†—
๐Ÿ“ Bangkok, Bangkok, ThailandFull time

About this role

As one of our early AI Engineering hires, you'll help define what AI at Cube looks like. You'll build the AI features people actually use from our self-hosted chat interface and MCP server to retrieval pipelines, prompts, evaluations, and integrations with internal systems. You'll work closely with our Infrastructure and Data Engineering teams to design architecture, connect systems, and transform emerging AI capabilities into practical products and tools that solve real problems every day.

  • Maintain and tunning our self-hosted chat interface including model connections, MCP integration, RAG/knowledge base setup
  • Build the RAG pipeline: ingestion, chunking, embeddings, vector store, retrieval, reranking, and evaluation
  • Integrate LiteLLM or OpenRouter as the gateway; handle routing, fallbacks, rate limits, and cost tracking
  • Maintain and configure MCP server and the tools it exposes to the model
  • Write prompts and evaluations, and iterate on them based on real usage and failure cases
  • Monitoring the logging, tracing, and guardrails of our AI platforms and model does.
  • Good to have exposure on MLOps/Platform team to deploy self-hosted models (vLLM, TGI, Ollama) and keep them healthy
  • Ship features end-to-end: API, retrieval, prompt, evaluation, and rollout
  • 4+ years of software engineering experience
  • Familiarity with containerized technologies and orchestration platforms such as Kubernetes
  • Strong interest in AI, LLMs, and the rapidly evolving model ecosystem
  • 1+ years of experience building, deploying, or supporting production LLM systems (RAG, agents, or fine-tuned models)
  • Experience deploying and configuring self-hosted LLM chat interfaces (Open WebUI preferred; similar platforms are acceptable)
  • Hands-on experience with retrieval and RAG systems, including embeddings, vector databases, chunking strategies, hybrid search, and evaluation methodologies
  • Experience working with LLM gateways or routing layers such as LiteLLM, OpenRouter, Portkey, or similar solutions
  • Experience serving open-weight models using tools such as vLLM, TGI, or SGLang
  • Experience designing and implementing secure integrations between LLMs and internal business systems
  • Nice to have: Experience with or understanding of MCP servers, agent frameworks, or tool-calling architectures
  • Nice to have: Experience with or understanding of LLM observability and monitoring platforms such as LangSmith, Langfuse, or similar tools

Frequently Asked Questions

Is the salary disclosed for the AI Engineer position at Cube?
The salary for this AI Engineer role at Cube is not publicly listed. Click "Apply Now" to learn more about the compensation package on their official careers page.
Where is the AI Engineer position at Cube located?
This AI Engineer role at Cube is based in Bangkok, Bangkok, Thailand. The position is listed as on-site or hybrid. Check the full job description or apply directly to confirm the work arrangement.
Is the AI Engineer role at Cube full-time or part-time?
This is listed as a Full time position. It is posted as a AI Engineer role in the Data, Product & Technology department at Cube.
Which team or department does the AI Engineer at Cube belong to?
This AI Engineer position is part of the Data, Product & Technology department at Cube. See the full job description for more information about the team structure and responsibilities.
How do I apply for the AI Engineer position at Cube?
Click the "Apply Now" button on this page. You will be redirected to Cube's official application portal hosted on workable where you can submit your application directly.
When was the AI Engineer job at Cube posted?
This AI Engineer position at Cube was posted on Jun 24, 2026. Apply as soon as possible โ€” early applications are often reviewed first.

You'll be redirected to Cube's official application page on workable.