Services
Contact Us
Şevval Alper

Şevval Alper

AI Researcher
21 Articles
Stay up-to-date on B2B Tech
Şevval is an AI researcher at AIMultiple. She has previous research experience in pseudorandom number generation using chaotic systems.

Research interests

Şevval focuses on AI coding tools, AI agents, and quantum technologies.

She is part of the AIMultiple benchmark team, conducting assessments and providing insights to help readers understand various emerging technologies and their applications.

Professional experience

She contributed to organizing and guiding participants in three “CERN International Masterclasses - hands-on particle physics” events in Türkiye, working alongside faculty to facilitate learning.

Education

Şevval holds a Bachelor's degree in Physics from Middle East Technical University.

Latest Articles from Şevval

AI
Benchmark
Aug 21

Speech-to-Text Benchmark: Deepgram vs. Whisper

We benchmarked the leading speech-to-text (STT) providers, focusing specifically on healthcare applications. Our benchmark used real-world examples to assess transcription accuracy in medical contexts, where precision is crucial. Based on both word error rate (WER) and character error rate (CER) results, GPT-4o-transcribe demonstrates the highest transcription accuracy among all evaluated speech-to-text systems. Deepgram Nova-v3 and…

AI
Benchmark
Aug 21

OCR Benchmark: Text Extraction / Capture Accuracy

OCR accuracy is critical for many document processing tasks, and SOTA multi-modal LLMs are now offering an alternative to OCR. We benchmarked leading OCR services in DeltOCR Bench to identify their accuracy levels in different document types: The full names of the above products and their versions in use as of November 2025 are listed…

Enterprise Software
Insight
Aug 21

Quantum Annealing: Practical Quantum Computing

Quantum annealing is a promising quantum technology for companies with urgent optimization problems that traditional computers cannot solve quickly. It can be used to solve optimization problems more effectively than traditional computers. However, it is still mostly used in academia, and more R&D is required to build commercial quantum annealers. There are different approaches to…

AI
Benchmark
Aug 14

AI Code Review Tools Benchmark

With the increased use of AI coding tools, codebases have become more prone to vulnerabilities, which increased the need for effective code reviews. To address this, we introduce RevEval (AI Code Review Eval), which benchmarks the top four AI code review tools across 309 pull requests from repositories of varying sizes and evaluates their performance…

Agentic AI
Benchmark
Aug 11

Top Agent Harnesses: Claude Code vs Codex

Agent harnesses serve as the production runtime for AI agents, with design choices that create performance variation across identical underlying models. We benchmarked 17 agent harnesses across 10 coding tasks. To isolate the harness rather than the model, we ran every agentic CLI on a single foundation model, Claude Sonnet 4.6 (non-reasoning), and AI code…

AI
Benchmark
Aug 11

Agentic IT: Can LLMs Design a Benchmark

We gave 12 large language models the job a benchmark team does: invent a benchmark, build it, run four models through it, and report the results. Each did it twice. None of the 24 attempts passed every criterion, and six of the rubric’s checks were passed by none of them. The two topics are text-to-SQL,…

Agentic AI
Benchmark
Aug 10

VELC-Bench: Verification on Long Context Benchmark

The model’s ability to locate a specific metric in context, compare its value to a claim, and confirm or reject it. This tests fine-grained value matching under long-context conditions. The model must both retrieve the value and perform a precise comparison. The models are tested in the following context windows: claude-fable-5 scores 90.0% on verify…

Enterprise Software
Open World Evaluation
Jul 30

MSP Automation: Acronis, ConnectWise Automate & Rewst

Managed service providers (MSPs) handle a constant operational load, including ticket management, patch management, onboarding, alert monitoring, billing reconciliation, and documentation updates. These are necessary but time-intensive tasks. Automation changes the equation by reducing manual workload and human error risk, enabling proactive responses through continuous system monitoring, and improving response times and consistency across client…

AI
Open World Evaluation
Jul 29

Top 8 Open Source AI Coding Agents

In prior evaluations, we benchmarked both open-source and proprietary Agentic CLIs, focusing on their performance in web development tasks, and some open-source agents performed as successfully as the paid options. Therefore, we also listed the top open-source coding agents for users with privacy concerns. For methodology, see the AI coding benchmark. For more details about…

Enterprise Software
Feature Comparison
Jul 14

Top 10 Google Colab Alternatives

Google Colaboratory is a popular platform for data scientists and machine learning scientists, but its limitations and pricing may not meet your needs. Several alternatives offer unique features and capabilities that cater to different data science needs and scenarios. Follow the links to see the top Google Colab alternatives: Cloud-based platforms offer scalable and flexible…