SIGN IN
Remote | Computer Vision Research & Evaluation Specialist — $75–$105/hour jobs in United States
info-icon
This job has closed.
company-logo

24-MAG · 3 weeks ago

Remote | Computer Vision Research & Evaluation Specialist — $75–$105/hour

24-MAG is offering a part-time consulting opportunity for computer vision professionals in the United States. The role involves designing and evaluating advanced AI models and tasks, focusing on real-world applications and performance analysis in computer vision.
Management Consulting
badNo H1Bnote

Responsibilities

Design challenging computer vision problems based on practical industry or research experience
Create tasks involving object detection, segmentation, recognition, tracking, generation, or multimodal reasoning
Target specific reasoning and capability gaps in advanced vision models
Define clear specifications, expected behaviours, constraints, and evaluation criteria
Develop accurate reference solutions and supporting materials using Python
Integrate tasks into agent-based development and evaluation environments
Create executable tests, validation scripts, datasets, or scoring logic where appropriate
Ensure reference implementations are technically sound, reproducible, and appropriately challenging
Evaluate model and agent performance across assigned computer vision tasks
Compare generated outputs with reference solutions and expected results
Identify tasks where models demonstrate meaningful limitations or inconsistent behaviour
Classify failures involving perception, localisation, visual reasoning, instruction following, or multimodal understanding
Document findings through clear and technically detailed written analysis
Review tasks and evaluation methods developed by other computer vision specialists
Maintain consistent standards for difficulty, accuracy, realism, and technical quality
Participate in calibration and peer-review activities
Collaborate with other subject matter experts to improve evaluation coverage and reliability

Qualification

Computer VisionDeep LearningPythonMultimodal SystemsModel EvaluationFailure AnalysisObject DetectionSegmentationRecognitionVision-Language ReasoningImage GenerationVideo GenerationPyTorchTensorFlowMachine Learning FrameworksTechnical Problem DesignReference Solution DevelopmentExecutable Test DevelopmentValidation Script DevelopmentDataset CreationScoring Logic DevelopmentTechnical Explanation

Required

Hands-on experience in applied research, deep learning, Python, multimodal systems, and evaluation of advanced AI models
Ability to design challenging computer vision problems based on practical industry or research experience
Create tasks involving object detection, segmentation, recognition, tracking, generation, or multimodal reasoning
Define clear specifications, expected behaviours, constraints, and evaluation criteria
Develop accurate reference solutions and supporting materials using Python
Integrate tasks into agent-based development and evaluation environments
Create executable tests, validation scripts, datasets, or scoring logic where appropriate
Ensure reference implementations are technically sound, reproducible, and appropriately challenging
Evaluate model and agent performance across assigned computer vision tasks
Compare generated outputs with reference solutions and expected results
Identify tasks where models demonstrate meaningful limitations or inconsistent behaviour
Classify failures involving perception, localisation, visual reasoning, instruction following, or multimodal understanding
Document findings through clear and technically detailed written analysis
Review tasks and evaluation methods developed by other computer vision specialists
Maintain consistent standards for difficulty, accuracy, realism, and technical quality
Participate in calibration and peer-review activities
Collaborate with other subject matter experts to improve evaluation coverage and reliability
Deep hands-on experience in computer vision through applied industry work, research, or graduate-level study
Practical proficiency in Python demonstrated through professional, academic, or open-source work
Strong understanding of modern computer vision methods, deep learning architectures, and evaluation techniques
Experience with PyTorch, TensorFlow, or comparable machine learning frameworks
Ability to design realistic technical problems and develop complete reference solutions
Strong written communication and the ability to explain complex model behaviour clearly
Ability to work independently and manage technical assignments effectively
Reliable availability for approximately 20 hours per week
A degree in computer science, machine learning, electrical engineering, robotics, applied mathematics, or a related technical field is highly relevant
Graduate or doctoral research in computer vision, machine learning, artificial intelligence, or multimodal systems may be especially valuable
Equivalent professional experience in applied computer vision may also be considered
Published research, open-source contributions, or production computer vision work may strengthen an application
Working proficiency in Python is required

Preferred

Experience with object detection, semantic or instance segmentation, image classification, or visual recognition
Background in vision-language models, multimodal reasoning, visual question answering, or image-text alignment
Experience with image or video generation, diffusion models, or generative vision systems
Familiarity with benchmark design, model evaluation, error analysis, or adversarial testing
Experience creating executable technical assessments or automated evaluation pipelines
Previous involvement in AI training, model evaluation, data annotation, or quality-review programmes
Familiarity with agent-based development environments and structured technical rubrics

Benefits

Part-time W-2 contingent employment arrangement
Fully remote role available to candidates based in the United States
Competitive rates between $75–$105 per hour depending on expertise and project scope

Company

24-MAG

linkedin
company-logo
At 24-MAG, we support emerging AI and consulting platforms by sourcing and connecting qualified professionals with remote, contract-based opportunities.

Funding

Current Stage
Early Stage
Company data provided by crunchbase