Skip to content
View GioiaZheng's full-sized avatar
💚
💚

Block or report GioiaZheng

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
GioiaZheng/README.md

Gioia Zheng

B.Sc. Student in Applied Computer Science and Artificial Intelligence
Sapienza University of Rome, Italy

Website · Hugging Face · LinkedIn · Email · Academic CV


About Me

I study information retrieval and retrieval-augmented generation, focusing on when improvements in retrieval do—or do not—lead to more accurate and grounded answers.

My work combines controlled retrieval and generation experiments with paired evaluation, failure analysis, versioned manifests, and inspectable artifacts.


Selected Work

Project Focus Current scope
msmarco-genqa Retrieval-augmented generation on MS MARCO Retrieval, reranking, generation, grounding analysis, paired statistical evaluation, and reproducible experiment reports
rag-observatory · Live Space · Toy Dataset Trace-based analysis for RAG systems Research prototype for inspecting retrieved evidence, generated answers, execution traces, and failure labels
CiboCompass Offline-resilient mobile food exploration React Native, Go, and SQLite application with local caching, a persistent retry queue, and idempotent submission

Current Research

Research question: When does better retrieval improve grounded generation, and when do conventional evaluation metrics hide the failure?

Current study: Controlled retrieve–rerank–generate experiments on MS MARCO, using paired statistical evaluation, explicit grounding measures, and per-example failure analysis.

Research direction: Developing a versioned RAG failure taxonomy that distinguishes retrieval, reranking, evidence-use, and generation errors.

Pinned Loading

  1. msmarco-genqa msmarco-genqa Public

    RAG-based question answering system on MS MARCO with retrieval, reranking, evaluation, and reproducibility checks.

    Python 18

  2. rag-observatory rag-observatory Public

    Research prototype for trace-based observability and failure analysis in retrieval-augmented generation.

    Python 2

  3. CiboCompass CiboCompass Public

    Mobile food exploration app with offline-resilient rating delivery, React Native, Go APIs, and SQLite.

    JavaScript 11

  4. handwritten-ocr-system handwritten-ocr-system Public

    Handwritten text recognition system using CNN-RNN-CTC models with CER/WER evaluation and reproducible training checks.

    Python 10