远程工作雷达

AI高级工程师(视觉)

AI Senior Engineer (Vision)

AI开发工程限定地区(需当地身份)
公司Able
薪资未公开
工作地点LATAM
地域资格限定地区(需当地身份)
时区要求无特别要求
用工类型Full-Time
发布时间今天
数据来源Jobicy
前往 Jobicy 查看并投递 →
注意地域限制:该职位明确限定在 LATAM 招聘。如果你是位于中国大陆的求职者,通常需要当地工作身份才能投递,或需与雇主确认是否接受独立合同(Contractor)形式合作。

我们从2012年开始,是一群工程师和设计师,决定要创造一些东西——于是我们做到了。Able最初是一个工程和产品中心,为一系列早期初创公司开发产品。我们在开发那些经过深思熟虑、高效且真正有用的产品过程中建立了许多关系。但自那以后,我们不断成长……我们的抱负也在增长。

现在,我们进入新的篇章——以应用人工智能为核心。AI是端到端软件开发周期中的强大推动力,我们正在建立实践方法,使我们比传统方法更快、更有效地交付软件,为我们的合作伙伴创造有意义的价值。如今,我们的建造者思维正推动我们成为每个职能都具备人工智能原生能力的组织。我们仍在进化,这也是机会的一部分。如果你希望与一支有雄心的团队一起创造、学习并迎接挑战,让我们一起创造吧。

这个职位在拉美地区100%远程办公。

你将负责的工作

我们正在寻找一位喜欢在计算机视觉与逻辑交汇处工作的人员。你将负责系统的“眼睛”和“大脑”——从视觉文档中提取复杂数据,然后协调这些数据如何被大型语言模型使用。

简而言之,我们希望找到这样的人:

  • 解锁视觉数据:构建能够“阅读”复杂文档的流程,使用视觉-语言模型(GPT-4V、Claude 3.5)和版面分析来理解布局、图表和视觉上下文。
  • 协调智能:负责应用逻辑层。你将使用LangChain或LangGraph来构建代理和链,查询我们的数据,对其进行推理并生成响应。
  • 原生PDF处理:处理PDF处理的混乱现实(PyMuPDF、版面解析),在AI看到之前保留结构。
  • 提示工程与逻辑:设计复杂的提示和控制流程,确保模型准确解读财务图表和布局,避免幻觉。
  • 成本与扩展性:应用成本优化思维(批量处理、模型选择),确保我们的视觉和协调层在经济上可行。

我们寻找的人

我们希望与那些热爱与团队协作、在构建软件的同时与同事建立包容和尊重关系的人共事。我们希望与那些坦诚面对自身不足和当前不了解的事物,但仍保持持续成长和弥补差距热情的人合作。

查看英文原文

Back in 2012, we were a group of engineers and designers who decided we wanted to build things—so we did. Able started as an engineering and product hub building for a portfolio of early-stage startups. We built many relationships while developing products that were thoughtful, effective, and genuinely useful. But, since then, we’ve grown… and so has our ambition.

Now, we’re entering our next chapter—defined by applied AI. AI is a powerful force in the end-to-end software development cycle, and we’re creating practices that allow us to deliver software fast and more effectively than traditional approaches, creating meaningful value for our partners. Today, our builder mindset is driving us to become an AI-native organization across every function. We’re still evolving, and that’s part of the opportunity. If you want to build, learn, and tackle challenges alongside an ambitious team, let’s build together.

This position is 100% remote within LatAm.

What you’ll be doing

We are seeking someone who enjoys working at the cutting edge where Computer Vision meets Logic. You will be responsible for the "eyes" and the "brain" of our system—extracting complex data from visual documents and then orchestrating how that data is used by Large Language Models.

In short, someone who likes:

  • Unlocking Visual Data: Building pipelines that can "read" complex documents, understanding layout, charts, and visual context using Vision-Language Models (GPT-4V, Claude 3.5) and Layout Analysis.
  • Orchestrating Intelligence: Owning the application logic layer. You will use LangChain or LangGraph to build the agents and chains that query our data, reason about it, and generate responses.
  • Native PDF Handling: Handling the messy reality of PDF processing (PyMuPDF, layout parsing) to preserve structure before the AI even sees it.
  • Prompt Engineering & Logic: Crafting complex prompts and control flows to ensure models interpret financial charts and layouts accurately without hallucinating.
  • Cost & Scale: Applying a cost-optimization mindset (batch processing, model selection) to ensure our vision and orchestration layers are economically viable.

What we’re looking for

We want to work with people who have a passion for collaborating with their teams, building software while nurturing inclusive and respectful relationships with their coworkers. With the ones that are open about their shortcomings and what they do not know now, but remain eager to keep on growing and closing those gaps.

Ideally, they would also have:

  • LLM Orchestration (Must Have): Deep experience with LangChain, LangGraph, or similar frameworks. You know how to manage context windows, tool calling, and agentic workflows.
  • Multimodal AI Experience: Hands-on experience integrating state-of-the-art vision models (GPT-4V, Claude 3.5 Sonnet) and embedding models (CLIP).
  • Document Intelligence Specialist: Familiarity with specialized models (e.g., Donut, Pix2Struct) and tools like Unstructured.io or Docling.
  • PDF Processing Mastery: Mastery over tools like PyMuPDF or pdfplumber for native element extraction.
  • Python ML Stack: Strong proficiency in PyTorch or TensorFlow.

Nice-to-Have:

  • Fine-Tuning: Experience fine-tuning vision or language models, specifically to improve accuracy on domain-specific artifacts like financial charts or tables.
  • Domain Knowledge: Prior experience handling documents in the Real Estate or Finance sectors.

Able is powered by curious, thoughtful people who care about what they build and how they build it. We’re actively investing in our team through AI training, knowledge-sharing, and hands-on experimentation to ensure everyone grows alongside the technology.

This position is 100% remote within LatAm. Strong verbal and written communication skills in English are a requirement. As a team member, you can expect:

  • To work 40 hours per week, and be available during normal business hours as needed.
  • Payments made in USD.
  • 18 days of PTO per year, observance of local holidays, and an annual break between Christmas and New Years.
  • Wellness + Remote Stipend
  • AI Voucher

About Able

Able builds technology products in a portfolio model. We believe that people, teams, and processes are more important than the ideas themselves, so we’ve focused on bringing great people together, and investing in their growth.

We’ve built products in a variety of industries. Everything from media to finance to toys to healthcare. Sometimes we work with management teams to help their businesses grow faster or unlock value using technology. Other times we start or buy businesses outright. Each time, we look for opportunities to leverage technology built at the portfolio-level to drive value faster.

Able is committed to inclusion and diversity and is an equal-opportunity employer. All applicants will receive consideration without regard to race, color, religion, gender, gender identity, sexual orientation, national origin, disability, or veteran status.

This is but the beginning of a conversation we’d love to have with you.

Apply, and let’s get this adventure started!

本页面信息整理自 Jobicy,版权归原发布方所有。职位可能随时关闭,投递请以原始页面为准。 本站只做信息聚合展示,不参与招聘流程,也不向求职者收取任何费用。

该公司其他在招职位

← 返回全部职位