Senior software engineer evaluating Codex, Claude Code, and Cursor interactions for G2i’s engineering team. Providing rigorous feedback on AI reasoning, code quality, and developer trust.
Responsibilities
Evaluate the quality of interactions with modern coding agents such as OpenAI Codex and Claude Code
Assess whether AI-generated responses make sense and whether preambles and reasoning are useful
Staff Software Engineer owning OAuth, authorization, and agent delegation systems. Building governed identity infrastructure for Redpanda’s enterprise AI data platform.
Senior software engineer evaluating AI coding agents for G2i’s engineering team. Assessing reasoning, explanations, and engineering judgment in Codex, Claude Code, and Cursor interactions.
Senior software engineer evaluating AI coding agents such as Codex, Claude Code, and Cursor. Providing rigorous written and video feedback on engineering quality.
Senior software engineer evaluating Codex, Claude Code, and Cursor interactions. Providing rigorous written and video feedback on AI - generated coding quality for G2i.
Senior engineer evaluating Codex, Claude Code, and Cursor interactions for G2i. Providing rigorous written and video feedback on AI - generated engineering work.
Applied AI Lead deploying Heidi’s healthcare AI for strategic healthcare accounts. Driving adoption, expansion, renewal, and measurable clinical and operational value.
Software Engineer, AI building secure, production - ready LLM and cloud - native solutions for Softchoice, an IT solutions provider. Delivering RAG, APIs, automation, and enterprise AI integrations.
Full - stack developer building TypeScript AI customer - support agents for gaiia’s telco operating system. Developing agent runtimes, omnichannel integrations, evaluation tooling, and safety guardrails.
Senior AI Engineer embedding with clinical, product, and operations teams. Shipping validated, monitored AI features for Prenuvo’s proactive whole - body healthcare platform.
Backend Software Engineer creating realistic coding challenges, bug fixes, and deterministic verifiers for Gramian’s AI training projects. Evaluating backend code quality, reliability, testing, and performance.