Services
Contact Us
Şevval Alper

Şevval Alper

AI Researcher
19 Articles
Stay up-to-date on B2B Tech
Şevval is an AI researcher at AIMultiple. She has previous research experience in pseudorandom number generation using chaotic systems.

Research interests

Şevval focuses on AI coding tools, AI agents, and quantum technologies.

She is part of the AIMultiple benchmark team, conducting assessments and providing insights to help readers understand various emerging technologies and their applications.

Professional experience

She contributed to organizing and guiding participants in three “CERN International Masterclasses - hands-on particle physics” events in Türkiye, working alongside faculty to facilitate learning.

Education

Şevval holds a Bachelor's degree in Physics from Middle East Technical University.

Latest Articles from Şevval

AI
Benchmark
Aug 2

AI Coding Benchmark: Claude Code vs Cursor

In AI coding, the market has fragmented into two categories: Agentic CLI tools and AI code editors embedded in IDEs. Each claims to automate development. Few comparisons show how they differ under identical workloads. We benchmarked each agent across 10 full-stack web development tasks, performing ~600 atomic validation checks per agent and more than 9,600…

Enterprise Software
Open World Evaluation
Jul 30

MSP Automation: Acronis, ConnectWise Automate & Rewst

Managed service providers (MSPs) handle a constant operational load, including ticket management, patch management, onboarding, alert monitoring, billing reconciliation, and documentation updates. These are necessary but time-intensive tasks. Automation changes the equation by reducing manual workload and human error risk, enabling proactive responses through continuous system monitoring, and improving response times and consistency across client…

AI
Open World Evaluation
Jul 29

Top 8 Open Source AI Coding Agents

In prior evaluations, we benchmarked both open-source and proprietary Agentic CLIs, focusing on their performance in web development tasks, and some open-source agents performed as successfully as the paid options. Therefore, we also listed the top open-source coding agents for users with privacy concerns. For methodology, see the AI coding benchmark. For more details about…

Enterprise Software
Feature Comparison
Jul 14

Top 10 Google Colab Alternatives

Google Colaboratory is a popular platform for data scientists and machine learning scientists, but its limitations and pricing may not meet your needs. Several alternatives offer unique features and capabilities that cater to different data science needs and scenarios. Follow the links to see the top Google Colab alternatives: Cloud-based platforms offer scalable and flexible…

Agentic AI
Benchmark
Jul 9

Top 4 AI Search Engines Compared

Searching with LLMs has become a major alternative to Google search. We benchmarked the following AI search engines to see which one provides the most correct results: Deepseek is the leader of this benchmark, by correctly providing 57% of the data in our ground truth dataset. You can also read our AI deep research benchmark…

AI
Benchmark
Jul 2

Speech-to-Text Benchmark: Deepgram vs. Whisper

We benchmarked the leading speech-to-text (STT) providers, focusing specifically on healthcare applications. Our benchmark used real-world examples to assess transcription accuracy in medical contexts, where precision is crucial. Based on both word error rate (WER) and character error rate (CER) results, GPT-4o-transcribe demonstrates the highest transcription accuracy among all evaluated speech-to-text systems. Deepgram Nova-v3 and…

Agentic AI
Benchmark
Jul 1

MCP Benchmark: Top MCP Servers for Web Access

We benchmarked 8 MCP servers across web search and extraction, as well as browser automation tasks, by running 4 different tasks 5 times on all suitable MCPs. We also performed a load test involving 250 concurrent AI agents. *Web search & extraction tasks are run with Bright Data’s default MCP server, browser automation tasks are…

AI
Insight
Jun 26

LLM Parameters: GPT-5 High, Medium, Low and Minimal

Some LLMs, such as OpenAI’s GPT-5 family, come in different versions (e.g., GPT-5, GPT-5-mini, and GPT-5-nano) and with various parameter settings, including high, medium, low, and minimal. Below, we explore the differences between these model versions by gathering their benchmark performance and the costs to run the benchmarks. We used the GPT-5 family in our…

AI
Benchmark
Jun 25

E-Commerce AI Video Maker Benchmark: Veo 3 vs Kling

Product visualization plays a crucial role in e-commerce success, yet creating high-quality product videos remains a significant challenge. Recent advancements in AI video generation technology offer promising solutions. We compared the top 6 AI video makers using 12 image-and-prompt inputs to evaluate their capabilities in generating product demonstration videos: Check out our methodology and evaluation…