Researcher Evidence Record
Tadashi Kozuno
This bounded record lists 29 Pith paper rows and 0 imported work rows attributed to this corpus identity. The enumerated, non-disputed paper rows include cs.LG, stat.ML, cs.RO work dated 2017 to 2026. The record describes sources and coverage; it makes no judgment about the person.
Compiled coverage vector
A sourced case file for attributed work. It is neither a profile score nor a verdict about this researcher.
Attributed works
A bounded ledger from the Pith paper and imported-work queries. Counts and source confidence stay with each work.
-
2026 Pith paper
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 4
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: a current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2026 Pith paper
The Harder Path: Last Iterate Convergence for Uncoupled Learning in Zero-Sum Games with Bandit Feedback
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- backfill
- Printed name
- Tadashi Kozuno
- Author position
- 3
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: a current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2026 Pith paper
Optimal last-iterate convergence in matrix games with bandit feedback using the log-barrier
paper citation record paper evidence challenge this paper
Sources and evidence
- Authorship source
- backfill
- Printed name
- Tadashi Kozuno
- Author position
- 3
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: a current Pith review exists.
- Citation counts
-
- 1 pith inbound references from cited_work_pith_inbound_counts
-
2026 Pith paper
Am I More Pointwise or Pairwise? Revealing Position Bias in Rubric-Based LLM-as-a-Judge
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 3
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: a current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2025 Pith paper
MK2 at PBIG Competition: A Prompt Generation Solution
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 5
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2024 Pith paper
Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form
paper citation record paper evidence challenge this paper
Sources and evidence
- Authorship source
- backfill
- Printed name
- Tadashi Kozuno
- Author position
- 2
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: a current Pith review exists.
- Citation counts
-
- 2 pith inbound references from cited_work_pith_inbound_counts
-
2024 Pith paper
Symmetry-aware Reinforcement Learning for Robotic Assembly under Partial Observability with a Soft Wrist
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 2
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2024 Pith paper
A Policy Gradient Primal-Dual Algorithm for Constrained MDPs with Uniform PAC Guarantees
paper citation record paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 2
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
-
- 2 pith inbound references from cited_work_pith_inbound_counts
-
2023 Pith paper
Multi-Agent Behavior Retrieval: Retrieval-Augmented Policy Training for Cooperative Push Manipulation by Mobile Robots
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 3
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2023 Pith paper
Local and adaptive mirror descents in extensive-form games
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 3
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2023 Pith paper
DoMo-AC: Doubly Multi-step Off-policy Actor-Critic Algorithm
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 2
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2023 Pith paper
Regularization and Variance-Weighted Regression Achieves Minimax Optimality in Linear MDPs: Theory and Practice
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 2
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2023 Pith paper
Counterfactual Fairness Filter for Fair-Delay Multi-Robot Navigation
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 4
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2023 Pith paper
When to Replan? An Adaptive Replanning Strategy for Autonomous Navigation using Deep Reinforcement Learning
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 4
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2023 Pith paper
Benchmarking Actor-Critic Deep Reinforcement Learning Algorithms for Robotics Control with Action Constraints
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 3
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2023 Pith paper
Robust Markov Decision Processes without Model Estimation
paper citation record paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 3
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
-
- 2 pith inbound references from cited_work_pith_inbound_counts
-
2022 Pith paper
Adapting to game trees in zero-sum imperfect information games
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 3
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2022 Pith paper
Confident Approximate Policy Iteration for Efficient Local Planning in $q^\pi$-realizable MDPs
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 3
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2022 Pith paper
KL-Entropy-Regularized RL with a Generative Model is Minimax Optimal
paper citation record paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 1
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
-
- 2 pith inbound references from cited_work_pith_inbound_counts
-
2022 Pith paper
No More Pesky Hyperparameters: Offline Hyperparameter Tuning for RL
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 8
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2021 Pith paper
Greedification Operators for Policy Optimization: Investigating Forward and Reverse KL Divergences
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 4
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2021 Pith paper
Unifying Gradient Estimators for Meta-Reinforcement Learning via Off-Policy Evaluation
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 2
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2021 Pith paper
Model-Free Learning for Two-Player Zero-Sum Partially Observable Markov Games with Perfect Recall
paper citation record paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 1
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
-
- 1 pith inbound references from cited_work_pith_inbound_counts
-
2021 Pith paper
Co-Adaptation of Algorithmic and Implementational Innovations in Inference-based Deep Reinforcement Learning
paper citation record paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 2
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
-
- 1 pith inbound references from cited_work_pith_inbound_counts
-
2021 Pith paper
Policy Information Capacity: Information-Theoretic Measure for Task Complexity in Deep Reinforcement Learning
paper citation record paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 3
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
-
- 1 pith inbound references from cited_work_pith_inbound_counts
-
2021 Pith paper
Revisiting Peng's Q($\lambda$) for Modern Reinforcement Learning
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 1
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2020 Pith paper
Leverage the Average: an Analysis of KL Regularization in RL
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 2
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2019 Pith paper
Gap-Increasing Policy Evaluation for Efficient and Noise-Tolerant Reinforcement Learning
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 1
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
-
2017 Pith paper
Unifying Value Iteration, Advantage Learning, and Dynamic Policy Programming
paper paper evidence challenge this paper
Sources and evidence
- Authorship source
- arxiv_oai
- Printed name
- Tadashi Kozuno
- Author position
- 1
- Identity state
- provisional
- Source confidence
- 0.7
- Review coverage
- Measured: no current Pith review exists.
- Citation counts
- No source count is attached to this work row.
Evidence apparatus
The machinery behind this record. Every lane states whether Pith measured it, did not query it, could not reach it, or withheld it.
| Lane | State | Observed | Boundary and source |
|---|---|---|---|
| identity | Measured | 2 | Canonical identity row plus public typed identifiers. source=authors, author_identifiers |
| papers | Measured | 29 of 29 bounded rows | Rows attributed to this author UUID in the Pith corpus. source=paper_authors |
| works | Measured zero | 0 of 0 bounded rows | Imported works not duplicated by the paper ledger. source=author_works |
| reviews | Measured | 5 of 29 bounded rows | Coverage count only. No review outcome is projected onto the person. source=current_verdicts |
| citations | Measured | 8 of 29 bounded rows | Counts remain itemized by work and source. source=cited_works |
| coauthors | Measured | 50 of 29 bounded rows | Shared-work edges from admitted paper rows. source=paper_authors |
| account | Unavailable | No public count of 1 bounded rows | Account metadata is separate from corpus evidence. source=users.author_id |
Public identity sources
-
name variant
Tadashi Kozuno
Enumerated research scope
- cs.LG18 rows
- stat.ML4 rows
- cs.RO3 rows
- cs.CL2 rows
- cs.GT1 rows
- cs.MA1 rows
- 20171 rows
- 20191 rows
- 20201 rows
- 20216 rows
- 20224 rows
- 20238 rows
- 20243 rows
- 20251 rows
- 20264 rows
Record scope
The work queries are bounded. Missing rows may mean measured zero, an unavailable source, a query that did not run, or private data that Pith withheld. The lane table keeps those cases separate.
Paper findings remain attached to papers. They do not become findings about this researcher.
Self-published account annex
Linked Pith account
Unavailable No public Pith account is linked to this corpus identity.
The account lane is self-published. Linking proves account control only and changes no corpus fact.