feat: enhanced credibility scoring per spec Section 4.6
- CredibilityScorer: composite score with weighted components (source_base 30%, content_quality 25%, engagement_authenticity 20%, cross_source 15%, historical 10%) - score_source: base credibility from registry - score_content_quality: length, structure, metadata quality heuristics - score_engagement_authenticity: bot detection via engagement ratios (like/view, retweet/view, reply/view rates) - score_cross_source_corroboration: clustering by content similarity (Jaccard n-grams), unique source count in consensus cluster - _content_hash: MD5 normalization for deduplication - _text_similarity: Jaccard similarity on word trigrams - _cluster_by_similarity: clusters items by similarity to query text - All 46 core NLP tests pass
This commit is contained in:
@@ -1,13 +1,16 @@
|
|||||||
"""Output sinks"""
|
"""
|
||||||
|
Output Module — sinks for persistence and serving
|
||||||
from .hazelcast_sink import HazelcastSink
|
"""
|
||||||
from .clickhouse_sink import ClickHouseSink
|
from sentiment_engine.output.sinks import (
|
||||||
from .latticedb_sink import LatticeDBSink
|
SinkConfig,
|
||||||
from .manager import OutputManager
|
HazelcastSink,
|
||||||
|
ClickHouseSink,
|
||||||
|
OutputSinkManager,
|
||||||
|
)
|
||||||
|
|
||||||
__all__ = [
|
__all__ = [
|
||||||
|
"SinkConfig",
|
||||||
"HazelcastSink",
|
"HazelcastSink",
|
||||||
"ClickHouseSink",
|
"ClickHouseSink",
|
||||||
"LatticeDBSink",
|
"OutputSinkManager",
|
||||||
"OutputManager",
|
|
||||||
]
|
]
|
||||||
|
|||||||
Reference in New Issue
Block a user