Careers at Chirp Labs
Data Architect & Engineer
Knowledge graph and ingestion
- Location
- Sydney, Australia · hybrid
- Employment type
- Full-time, permanent · 38 hours a week
- Salary
- AUD 150,000 to 170,000 base, plus 12% superannuation and equity (ESOP)
- Languages
- English and Mandarin Chinese, including written Simplified Chinese
- Nominated occupation
- ANZSCO 261313 Software Engineer
- Applications close
- 20th of October, 2026
Chirp Labs is hiring a bilingual Data Architect & Engineer in Sydney to own the data layer behind Birdie: the knowledge graph and ingestion pipeline that turns sales calls, emails and CRM records into one verified record per deal. Australian citizens and permanent residents are encouraged to apply, and sponsorship is available under the Skills in Demand visa (subclass 482).
About Chirp
Chirp Labs builds Birdie, an AI-native voice agent that runs a sales rep's CRM. After every call the rep talks to Birdie, and Birdie updates the CRM, schedules follow-ups and saves context. Underneath Birdie sits a knowledge graph and ingestion pipeline that captures calls, emails and notes and resolves them into one clean, self-improving record per deal. That data layer is what this role owns.
We are a six-person startup founded in 2024, based in Sydney and San Francisco, backed by Startmate (2025 cohort) and angel investors, and CRM-agnostic by design. Our current markets are Oceania and North America, with European expansion planned.
Why this role exists
This role owns the data layer that Birdie acts on. Birdie only updates a CRM or schedules a follow-up from the resolved record of a deal, never from a raw transcript or a stale CRM field. Today that record is produced by a pipeline and knowledge graph the founders and Lead Engineer built over 18 months. As customers and CRMs multiply, it has to become a production data platform: one ontology, reliable ingestion from calls, email, calendar and every supported CRM, measurable data quality, and privacy controls a customer's security team will sign off on.
You will be the first dedicated data hire, setting the architecture and building most of it yourself. You work in English with the Sydney and San Francisco team, and in Chinese with our Chinese-speaking investors, partners and data sources.
What you will do
- Own the data architecture of Birdie's knowledge graph: the deal ontology, schemas, entity-resolution rules, naming conventions and data dictionary.
- Design, build and maintain ingestion pipelines that take calls, transcripts, emails, calendar events and CRM records (HubSpot, Salesforce, Pipedrive and others) and resolve them into one verified record per deal.
- Research, analyse and evaluate data and system needs with product and engineering; identify limits and deficiencies in the current pipelines and propose remedies.
- Test, debug, diagnose and correct faults; build data-quality checks, monitoring, backfills and replay so a bad extraction never reaches a customer's CRM.
- Design for privacy and security: personal-data handling, retention, per-customer isolation, access control, and compliance with the Australian Privacy Act and customer contracts.
- Write and maintain technical documentation, data contracts and operational runbooks for the data platform.
- Advise on platform design choices, including vendor evaluation, build-versus-buy and cost, and own the data platform's cloud spend.
- Work in Chinese and English: handle Chinese-language transcripts and CRM text in the pipeline (segmentation, entity resolution, evaluation), and produce technical documentation and data due-diligence answers for Chinese-speaking investors, partners and vendors.
What you need
Applicants must show all of the following.
- A bachelor's degree or higher in computer science, software engineering, data engineering or a related field, or equivalent professional experience.
- At least 3 years building production data pipelines and data models, with fluent Python and SQL.
- Experience designing schemas or ontologies and doing entity resolution, in a graph database or in relational models of graph-like data.
- Hands-on experience with a workflow or streaming tool (for example Airflow, Dagster, Inngest, Kafka or dbt) and a cloud data store (for example Postgres, BigQuery, Snowflake or object storage).
- Experience with LLM-based extraction: embeddings, retrieval, and evaluating the accuracy of structured data pulled from unstructured text.
- Working knowledge of data privacy and security practice: personal-data handling, access control, retention and tenant isolation.
- Professional working proficiency in English and Mandarin Chinese, including written Simplified Chinese: able to write technical documentation, review data and present designs in both.
- Australian work rights, or eligibility for sponsorship under the Skills in Demand visa (subclass 482).
Nice to have
None of these is required. They tell us you will be productive sooner.
- Early-stage startup experience, where you were the only or first data engineer.
- Familiarity with tools in our stack: Supabase and Postgres, Inngest, the Cognee knowledge-graph framework, Composio connectors, Deepgram, Vercel and Cloudflare.
- Experience with HubSpot, Salesforce or Pipedrive APIs and their data models.
- Chinese-language natural language processing: segmentation, named-entity recognition, or evaluating LLM output on Chinese text.
- TypeScript in a Next.js codebase, and daily use of AI coding tools such as Claude Code or Cursor, which is how this team writes most of its code.
- Exposure to SOC 2 or ISO 27001 controls for a SaaS product.
Why this role is bilingual
Bilingual Chinese and English is an inherent requirement of this position, not a preference, and it applies equally to every applicant, including Australian citizens and permanent residents.
- Investors and partners. Chirp produces its company and product materials in English and Simplified Chinese for Chinese-speaking investors and partners. This role answers their technical and data due-diligence questions directly, in Chinese, without a translator in the loop.
- Chinese-language data in the product. Birdie's pipeline processes sales calls, emails and CRM text. Where customers or their buyers work in Chinese, the person who designs segmentation, entity resolution and quality evaluation for that text has to read it.
- Vendors and engineering partners. Some of the data and AI vendors and engineering partners Chirp works with operate in Chinese; the role reviews their documentation, contracts and integration specs in the original.
Terms of employment
How to apply
Email nick@trychirp.com with the subject "Data Architect & Engineer". Include a CV in English (a Chinese CV is welcome as well), a short note in either language on a data system you designed and what you would change about it, and links to code you are proud of. Applications close on 20 October 2026. We reply to every applicant and interview in English and Chinese.