Skip to content
View jeknov's full-sized avatar
👩‍💻
👩‍💻

Block or report jeknov

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
jeknov/README.md

Hi there 👋

I'm Jekaterina Novikova, an AI researcher specializing in evaluation and reliability of foundation models - the science of knowing when language models can be trusted, and catching them when they can’t.

🧐 Research

  • 🤨 I’m currently working on improving trustworthiness and reliability of LLMs.
  • 🧠 Over the past several years, my team and I have studied the AI-based methods to detect cognitive impairment associated with dementia and mental illness from human language. This research has informed the development of practical healthcare applications used in clinical trials and senior care, improving the lives of people with dementia and psychiatric illness.
  • 🏆 With my postdoc colleagues I co-organized the End-to-End NLG shared task that attracted NLG researchers from across the world and set a new research agenda for the field of neural text generation.
  • 🤖 My research on conversational models for human-robot interaction was used for a multimodal dialogue system implemented in a humanoid robot engaging and interacting autonomously with customers in a grocery store in Scotland and in a busy shopping mall in Finland.

🌐 Socials

Scholar ResearchGate LinkedIn Twitter Website

💻 Skills

Python Shell Script PyTorch NumPy Pandas SciPy Plotly scikit-learn Huggingface Git Docker Trello

📔 Projects I have contributed to

Readme Card Readme Card

Pinned Loading

  1. awesome-llm-consistency awesome-llm-consistency Public

    A curated list of papers and resources about the consistency of large language models.

  2. google/BIG-bench google/BIG-bench Public archive

    Beyond the Imitation Game collaborative benchmark for measuring and extrapolating the capabilities of language models

    Python 3.2k 618

  3. EMNLP_17_submission EMNLP_17_submission Public

    The dataset and statistical analysis code released with the submission of EMNLP 2017 paper "Why We Need New Evaluation Metrics for NLG"

    R 19 3

  4. RankME RankME Public

    The dataset and code released with the submission of NAACL 2018 paper "RankME: Reliable Human Ratings for Natural Language Generation"

    HTML 25 5

  5. INLG_16_submission INLG_16_submission Public

    The dataset released with the submission of INLG 2016 paper "Crowd-sourcing NLG Data: Pictures Elicit Better Data" (https://aclweb.org/anthology/W/W16/W16-6644.pdf)

    6 1