Jina Reranker
Unknown
Jina Reranker is a neural reranking model that improves search and RAG systems by reordering retrieved documents for better query alignment.
A comprehensive index of artificial intelligence and machine-learning research with AI-generated summaries, citation metrics, and direct links to papers and code.
Unknown
Jina Reranker is a neural reranking model that improves search and RAG systems by reordering retrieved documents for better query alignment.
Unknown
Nvidia's Nemotron-4 340B models and reward model enable synthetic data generation for training smaller language models, with over 98% of alignment data being synthetic.
Unknown
Zephyr 7B uses distilled direct preference optimization (dDPO) and AI feedback data to improve intent alignment in chat-based language models.
Songyue Han, Mingyu Wang, Jialong Zhang, et al.
A comprehensive survey of LLMs covering architectures, key technologies, interdisciplinary integrations, optimization, applications, and challenges.
Haoru Tan, Wang Wang, Sitong Wu, et al.
Dataset Distillation by Influence Matching aligns the final outcome of training by learning a compact synthetic set whose effect on converged parameters matches that of the full dataset.
Yufei Wang, Wanjun Zhong, Liangyou Li, et al.
A comprehensive survey of alignment technologies for large language models, summarizing methods for effective high-quality alignment.
Unknown
This survey systematically reviews threats and countermeasures for trustworthy LLM agents, organizing defense approaches into three paradigms: alignment, monitoring, and control.
Ahmed Elnaggar, Michael Heinzinger, Christian Dallago, et al.
ProtTrans trains protein language models on up to 393 billion amino acids, showing raw embeddings capture biophysical features and outperform state-of-the-art in secondary structure prediction without evolutionary information.
Dwip Dalal, Shivansh Patel, Chahit Jain, et al.
Anchor-Align augments behavior cloning with vision-language anchoring and language-action alignment to prevent representation drift and improve VLA policy generalization.