Anyone AI
Anyone AI
Anyone AI ir uz Latīņameriku orientēta izglītības un talantu platforma, kas palīdz programmatūras inženieriem veidot karjeru mākslīgā intelekta un mašīnmācīšanās jomā. Tās attālinātajās, mentoru vadītajās programmās tehniskā apmācība, tostarp Machine Learning Developer virziens, tiek apvienota ar sagatavošanos karjerai, piemēram, CV un LinkedIn profila konsultācijām, interviju treniņu un angļu valodas praktizēšanu. Uzņēmums pārvalda arī talantu tirgu, kas savieno tā apmācīto kopienu ar darba devējiem, kuri meklē mākslīgā intelekta speciālistus, kā arī piedāvā uz darbā iekārtošanu vērstu modeli, kas ietver no ienākumiem atkarīgus maksājumus un partnerības ar darba devējiem.

Senior Software Engineer, SWE-Bench Evaluation

Review open-source software tasks from GitHub issues and pull requests for Anyone AI. Assess their clarity, test quality, reproducibility, complexity, and suitability for benchmarking.

Apraksts

  • Review and assess software engineering tasks based on real GitHub issues and pull requests.
  • Check that task descriptions and success criteria are clear and complete.
  • Evaluate unit tests for accuracy, coverage, and resilience.
  • Find flaky tests, missing dependencies, version mismatches, and environment problems.
  • Check whether tasks can be reproduced consistently across environments.
  • Judge each task’s practical difficulty and technical complexity.
  • Recommend whether tasks should be accepted, revised, or removed.
  • Review bug fixes and feature work, including repository setup, dependency handling, changes across files and modules, and overall technical quality.

Prasības

  • At least three years of professional software engineering experience.
  • Strong experience working in large codebases spanning multiple files.
  • Experience reviewing pull requests, debugging problems, and maintaining production software.
  • A solid understanding of unit testing and test coverage.
  • Ability to judge whether tests validate a solution correctly without unnecessarily limiting implementation choices.
  • Experience managing dependencies, configuring environments, and ensuring reproducibility.
  • Strong command of Git and GitHub-based development workflows.
  • Ability to analyze complex technical issues and provide clear written feedback.
  • Contributions to or maintenance of open-source projects are a plus.
  • Experience with SWE-Bench, SWE-Bench Verified, or comparable coding benchmarks is a plus.
  • Experience with major Python open-source projects, such as Django, Flask, scikit-learn, SymPy, matplotlib, requests, or pytest, is a plus.
  • Experience with Docker, CI/CD, pip, conda, or dependency pinning is a plus.
  • Knowledge of test fixtures, test isolation, or property-based testing is a plus.
  • Experience designing technical assessments or reviewing coding challenges is a plus.
  • Experience in AI/ML evaluation, data curation, RLHF, or benchmark development is a plus.

Priekšrocības

  • Work remotely.
  • Part-time, project-based consulting engagement.

Saistītās vakances