Skip to content
View achraf059's full-sized avatar
  • Chengdu , China
  • 05:43 (UTC -12:00)

Highlights

  • Pro

Block or report achraf059

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
achraf059/README.md

Hi, I'm Achraf Ait Tayeb 👋

Software Engineering student at Sichuan University, China, graduating in June 2027.

My current academic focus is AI for Software Engineering, especially LLM-based program repair, coding-agent reliability, and empirical evaluation. I also build systems projects to strengthen my understanding of distributed systems, fault tolerance, persistence, and software reliability.

I am currently preparing for Fall 2027 Master's applications.


🔬 Research

Coding-Agent Confidence Under Controlled Repository Evidence

Completed independent empirical Software Engineering study; manuscript in preparation.

I studied whether a coding agent can accurately judge, before attempting a repair:

  • how likely it is to successfully fix a repository-level software bug, and
  • whether the repository evidence it has been given is sufficient.

The final study used:

  • 60 real repository-level software bugs
  • 3 controlled evidence conditions
  • 3 replicates per task-condition
  • 540 repair attempts
  • executable Docker-based evaluation
  • a frozen experimental protocol and analysis plan
  • 10,000 task-level bootstrap resamples

Main finding

Relevant repository evidence substantially improved both confidence and actual repair performance:

  • predicted repair success increased by 55.1 percentage points
  • executable repair success increased by 20.0 percentage points

However, task-level confidence changes showed very little correspondence with task-level performance improvements (Pearson r = 0.073).

This suggests that coding models can recognize that relevant repository evidence is generally useful while still having limited ability to estimate how much that evidence will help on a specific software-repair task.

A public research artifact will be added after the manuscript and replication package are prepared.


🚀 Featured Projects

⚙️ Distributed Task Processor — Java 21

View repository

A coordinator/worker distributed task-processing system built from scratch in Java 21.

Engineering highlights

  • Custom length-prefixed JSON protocol over TCP
  • Single-writer coordinator event loop
  • Concurrent worker execution
  • Heartbeat-based worker failure detection
  • At-least-once execution semantics
  • Attempt leases and stale-result rejection
  • Automatic retry and reassignment
  • SQLite-backed durable state and coordinator restart recovery
  • Task deadlines and cooperative cancellation
  • Admission control and overload handling
  • Durable idempotent submissions
  • Persistent priority scheduling
  • Deterministic FIFO sequencing
  • Aging-based starvation prevention
  • GitHub Actions CI
  • 172 automated tests
    • 119 unit tests
    • 53 integration tests
    • 0 failures

Technologies: Java 21 · TCP · Concurrency · SQLite · Maven · JUnit · GitHub Actions


🎟️ Blaniko — Activity Discovery & Outing Planning

View repository

Casablanca-focused activity-discovery and outing-planning product built around structured, venue-verified data.

I lead product direction, venue-data design, QA, and an AI-assisted development workflow.

Current scope

  • structured dataset covering 99 verified venues
  • personalized outing-planning flows
  • activity discovery and filtering
  • favorites, collections, comparison, and saved outings
  • bilingual English/French support
  • venue verification and data-quality workflow
  • branch-and-PR development process

Technologies: React · TypeScript · Express · Supabase · PostgreSQL · Vite · GitHub Actions


📊 Yelp Big Data Analytics & Text-to-SQL

View repository

Originally developed as a company training project and later independently extended.

Work includes

  • Hadoop/HDFS distributed data processing
  • Hive/HiveQL analytics
  • Python and NLTK review-text analysis
  • Flask Text-to-SQL interface
  • schema-aware LLM prompting
  • SELECT-only SQL validation
  • automatic retry for failed queries
  • NOAA weather-data enrichment

Technologies: Python · Hadoop · HDFS · Hive · SQL · NLTK · Flask


🧰 Technical Stack

Programming: Python · Java · JavaScript · TypeScript · C · C++ · SQL

Systems & Data: TCP sockets · Concurrency · SQLite · Maven · JUnit · Docker · Linux · Hadoop · HDFS · Hive

Web & Tools: React · Express · Flask · Supabase · Git · GitHub Actions


🎓 Education

Sichuan University
Bachelor of Software Engineering
English-taught program
Expected graduation: June 2027


🌍 Languages

Arabic — Native
French — Fluent, DELF B2
English — Fluent, English-medium degree
Mandarin Chinese — Intermediate


📫 Contact

Pinned Loading

  1. blaniko blaniko Public

    Full-stack activity discovery and outing-planning platform for Casablanca with structured venue data, deterministic recommendations, an Express/Supabase backend, security hardening, and CI-tested w…

    TypeScript

  2. yelp-analysis-project yelp-analysis-project Public

    Big-data analytics platform using Hadoop, HDFS, Hive, Python, NLP, weather enrichment, and a schema-aware Text-to-SQL application.

    HTML

  3. distributed-task-processor distributed-task-processor Public

    Fault-tolerant distributed task processor in Java 21 with custom TCP messaging, concurrent workers, failure detection, retries, time-bounded leases, cooperative cancellation, bounded admission cont…

    Java