Skip to content
View K-Rishita's full-sized avatar

Block or report K-Rishita

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
K-Rishita/README.md

Hi there I'm Rishita πŸ‘‹

AI engineer and IEEE-published researcher with 2+ years of experience in LLM evaluation, enterprise AI deployment, and multi-agent systems. Currently pursuing my MS in Artificial Intelligence at Northeastern University (GPA: 3.91).


πŸ”¬ Research & Projects

πŸƒ The Long Con β€” Long-Horizon Social Deduction Benchmark for LLMs (Jan 2026 – Present)

Writing a preprint introducing a long-context, multi-turn benchmark for evaluating LLMs in adversarial multi-agent settings β€” measuring cooperation, deception, persuasion, and social reasoning under uncertainty. Preliminary results across 600+ games and 40+ models show that long-context interaction histories improve collaborative reasoning over compressed baselines.

🎭 Character Recognition in Live-Action vs. Animated Movies

Comparative study of face recognition models across live-action and animated video content. Evaluated multiple architectures trained from scratch on both domains, benchmarking recognition accuracy, detection rate, and domain gap across frame-level and video-level metrics. Findings highlighted how stylized animated features cause systematic failures in models designed for photorealistic faces.

🧭 Detection and Steering of LLM Outputs Using Backward Reachability (Oct – Dec 2025)

Extended BRT-Align to show that latent-space geometry governs the effectiveness of RL-based steering for correcting misaligned LLM outputs under adversarial attacks. Benchmarked transformer, CNN, and classical ML classifiers across diverse threat scenarios, demonstrating that high detection accuracy doesn't reliably translate to robust alignment.

⚑ Distributed Training Performance Analysis on GPU Clusters (Jan – Apr 2025)

Benchmarked single- and multi-node distributed training using PyTorch DDP with NCCL across CNNs, ResNet, and Vision Transformers on CIFAR-10. Quantified tradeoffs between batch size, gradient sync cost, and scaling efficiency across 1-GPU, 2-GPU, and multi-node configurations.

🫁 Hybrid CNN for Pulmonary Disease Detection β€” IEEE ICNC 2024 (Nov 2023 – Apr 2024)

Built a hybrid CNN model combining DenseNet121, ResNet50, and InceptionV3 to detect and classify multiple lung diseases from CT images β€” achieving 98.61% accuracy. Published at IEEE ICNC 2024.


πŸ›  Tech Stack

Python PyTorch TensorFlow LLMs RAG Multi-agent Systems Computer Vision NLP Distributed Training React Java MySQL


πŸ“« Let's connect

LinkedIn Email

Pinned Loading

  1. AshishT558/Genetic-Alg-Combat-Sim AshishT558/Genetic-Alg-Combat-Sim Public

    A combat simulator utilizing Genetic Algorithms

    Python 2

  2. AmongLLMs AmongLLMs Public

    Forked from haplesshero13/AmongLLMs

    Make LLM agents play the game "Among Us", and study how the models learn and express lying and deception in the game.

    Python 1

  3. haplesshero13/AmongLLMs haplesshero13/AmongLLMs Public

    Forked from 7vik/AmongUs

    Make LLM agents play the game "Among Us", and study how the models learn and express lying and deception in the game.

    Python 3 4

  4. Distributed-Training-CIFAR10 Distributed-Training-CIFAR10 Public

    Forked from vivekdhir77/Distributed-Training-CIFAR10

    Python

  5. MealMind MealMind Public

    Forked from TanviKandalla/MealMind

    TypeScript