Anyone AI
Anyone AI
Anyone AI — Lotin Amerikasiga yo‘naltirilgan ta’lim va iqtidorlar platformasi bo‘lib, dasturiy ta’minot muhandislariga sun’iy intellekt va mashinaviy o‘rganish sohalarida karyera rivojlantirishga yordam beradi. Uning masofaviy, mentorlar boshchiligidagi dasturlari texnik ta’limni, jumladan Machine Learning Developer yo‘nalishini, rezyume va LinkedIn bo‘yicha yo‘l-yo‘riq, suhbatga tayyorgarlik hamda ingliz tili amaliyoti kabi karyera tayyorgarligi bilan birlashtiradi. Kompaniya, shuningdek, o‘qitilgan hamjamiyatini sun’iy intellekt mutaxassislarini izlayotgan ish beruvchilar bilan bog‘laydigan iqtidorlar bozorini yuritadi. Bundan tashqari, daromaddan ulush asosidagi to‘lovlar va ish beruvchilar bilan hamkorlikni o‘z ichiga olgan, ishga joylashtirishga yo‘naltirilgan modelni taklif qiladi.

Senior Software Engineer, SWE-Bench Evaluation

Review open-source software tasks from GitHub issues and pull requests for Anyone AI. Assess their clarity, test quality, reproducibility, complexity, and suitability for benchmarking.

Tavsif

  • Review and assess software engineering tasks based on real GitHub issues and pull requests.
  • Check that task descriptions and success criteria are clear and complete.
  • Evaluate unit tests for accuracy, coverage, and resilience.
  • Find flaky tests, missing dependencies, version mismatches, and environment problems.
  • Check whether tasks can be reproduced consistently across environments.
  • Judge each task’s practical difficulty and technical complexity.
  • Recommend whether tasks should be accepted, revised, or removed.
  • Review bug fixes and feature work, including repository setup, dependency handling, changes across files and modules, and overall technical quality.

Talablar

  • At least three years of professional software engineering experience.
  • Strong experience working in large codebases spanning multiple files.
  • Experience reviewing pull requests, debugging problems, and maintaining production software.
  • A solid understanding of unit testing and test coverage.
  • Ability to judge whether tests validate a solution correctly without unnecessarily limiting implementation choices.
  • Experience managing dependencies, configuring environments, and ensuring reproducibility.
  • Strong command of Git and GitHub-based development workflows.
  • Ability to analyze complex technical issues and provide clear written feedback.
  • Contributions to or maintenance of open-source projects are a plus.
  • Experience with SWE-Bench, SWE-Bench Verified, or comparable coding benchmarks is a plus.
  • Experience with major Python open-source projects, such as Django, Flask, scikit-learn, SymPy, matplotlib, requests, or pytest, is a plus.
  • Experience with Docker, CI/CD, pip, conda, or dependency pinning is a plus.
  • Knowledge of test fixtures, test isolation, or property-based testing is a plus.
  • Experience designing technical assessments or reviewing coding challenges is a plus.
  • Experience in AI/ML evaluation, data curation, RLHF, or benchmark development is a plus.

Imtiyozlar

  • Work remotely.
  • Part-time, project-based consulting engagement.

O‘xshash ish o‘rinlari