Skip to content
View limorgu's full-sized avatar
  • United States

Block or report limorgu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
limorgu/README.md

Limor Kissos, PhD

AI Product / Program Manager | Independent AI Researcher | Human-Centered AI & Evaluation

PhD in Art Therapy with deep expertise in human behavior, emotional processing, and clinical research pipelines. Former Senior Program Manager / Knowledge Manager at Amazon (Middle Mile Transportation), scaling internal knowledge systems for 900+ people (250+ engineers, ~50 scientists).

Currently building modular AI evaluation frameworks, data pipelines from unstructured sources (books/memoirs/images), and metrics-driven tools for trustworthy, empathetic AI.


🛠 Featured AI Research Projects

AI-Powered Book Dataset Builder
Automates turning raw book page images into clean, structured, research-ready JSON datasets.

  • Three-stage pipeline: extraction → audit for gaps → precise gap-filling
  • Resume-aware, cost-efficient, and audit-driven
  • Ideal for large-scale qualitative research from physical books

Modular Multi-Stage AI Pipeline for Research Datasets
Full end-to-end framework for building structured datasets from physical books.

  • Stages include workspace setup, librarian intake (OCR), worker extraction, judge audit, analytics, reports, ground-truth export, and benchmarking
  • Highly configurable via JSON (domains, taxonomy, labels)
  • Supports deterministic/local and future LLM connectors

Raw Data → Meaningful Insights Pipeline
Stable, Codex-compatible version of the book processing workflow.

  • Turns raw inputs into categorized insights using configurable taxonomies
  • Includes full stage orchestration and reporting

MASK Honesty Benchmark Evaluation (Fork)
Implementation for evaluating AI honesty (disentangling it from accuracy) using the MASK benchmark.

  • Tests model consistency under pressure to lie

Other Projects


Professional Experience Highlights

  • Amazon (2021–2025): Senior Knowledge Manager / Program Manager
    Built and scaled data-driven knowledge & workflow systems for large ops teams. Designed metrics/experimentation frameworks, led AWS enablement workshops (900+ participants), science newsletter.

  • Independent Research (2023–Present):
    LLM evaluation (empathy, alignment, grounding, sycophancy via ELEPHANT-style studies), synthetic data generation, narrative/therapy AI tools, and memoir dysregulation analysis.

  • Academic Background: PhD Art Therapy (University of Haifa), published research on AI detection of childhood sexual abuse in drawings (~72% accuracy).


Skills & Tools

Product & Program Management — Strategy, experimentation (A/B), metrics & dashboards, cross-functional leadership, user research
AI/ML — Evaluation frameworks, RAG/pipelines, synthetic data, LLM prompting & evaluation
Technical — Python, Pandas, SQL, JSON data pipelines, Git, API integration



Last updated: June 2026

Popular repositories Loading

  1. mask mask Public

    Forked from centerforaisafety/mask

    Mask - Code for evaluating AI systems on the MASK honesty benchmark with GPT 3.5

    Python

  2. books_insights_project books_insights_project Public

    This tool automates turning raw book page images into a clean, structured, research-ready JSON dataset.

    Python

  3. nature-ai-pipeline nature-ai-pipeline Public

    Python

  4. pipeline_spec pipeline_spec Public

    This project builds a structured research dataset from physical books using a multi-stage AI pipelin

    Python

  5. comparing_architecture_classification- comparing_architecture_classification- Public

    Python

  6. from-raw-data-to-categories from-raw-data-to-categories Public

    This pipeline is aims to take raw books inputs and turn them into meaningful insights categories

    Python